Pith. sign in

Paper Citation Record · LEDGER

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2608.04436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04436 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:15:39.125513Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4cced5a-a747-4f11-ab60-922e101f0195 · outbound

This paper cites Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:15:41.386031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:37.003251Z digest=sha256:75a474774f91b92b668dcb51065ebe35e846fe7a7e34f4fc448cc72703a30600

Observation 589322b1-f649-457a-802b-72c2b985e498 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.175322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.175322Z digest=sha256:bf7a97367b9bac9e0969b8ee85d692c9006855eee835f765febf37d20afe73c7

Observation 304b4280-e528-4849-a48b-0aab005f9f19 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.290999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.290999Z digest=sha256:6a5a34f74512c5c0566b89d623080f2130baebe5b14151a9b5db49fe80d3c590

Observation bd365d23-98d6-4c9b-b397-2e92feba0156 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.407731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.407731Z digest=sha256:26860893782a40f2270045c5dd15978447da2ea1df7dd159f57aac6f338a4b76

Observation 653e64f5-e909-4828-89fa-a148fcc84429 · outbound

This paper cites GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.549104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.549104Z digest=sha256:3a963531752643692da009085bfad6beb2f382dad9f4a80a69ca69678ebebacc

Observation 0cbe4b8d-b2e4-4b47-8431-493a4cf4bce4 · outbound

This paper cites Re-Imagen: Retrieval-Augmented Text-to-Image Generator.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Re-Imagen: Retrieval-Augmented Text-to-Image Generator

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.706082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.706082Z digest=sha256:78d3f0e3cf1239c3f936092501602f14dc121be85b4b848e2b2e755171f0bfb1

Observation 045cae11-6dd4-45f1-ab2b-70bb179ab2f9 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.847345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.847345Z digest=sha256:2f11103018a639ae4682bf3f3dc520c86817715fdc061cd4bce955b5886037f3

Observation 3ae2b93c-2391-4649-8aea-bf48113677a2 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3.5: Native Multimodal Models are World Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.958795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.958795Z digest=sha256:204869d7ef7983d73d11a7bb34504448ac6e0dd05fcfdebb8d8ce776ed4db07c

Observation 3081dbd7-c498-4520-b610-47e6b2990731 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.098358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.098358Z digest=sha256:f35f7d43a75f6e7fdfbe4af41d78410114b7bf5a40ea2a157aa6ca71c110b06c

Observation 97a849ac-d1d9-43d2-beea-addf51bf98d6 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.217313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.217313Z digest=sha256:42badbbebef06fbaa67552cde1e64621fe5200a00bcd4009dfab2bc97477f684

Observation b8e6eff9-412a-43a9-9940-376db9b11c24 · outbound

This paper cites Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.364743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.364743Z digest=sha256:05a9fd77b7158a1cbdcb170eb8292c74d37f4701b232797abadee3b054b72fda

Observation 5da3c431-64e4-40f9-8d53-0f79d2030517 · outbound

This paper cites Demystifying Flux Architecture.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Demystifying Flux Architecture

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.504110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.504110Z digest=sha256:053255b52b7303ae4c20a62cb17b53a7c228ce0c460a197535b98c7e9e165c26

Observation fe808141-0b6a-4c12-9b3f-0087ebacd1b8 · outbound

This paper cites Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.622986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.622986Z digest=sha256:720eaa19c4807e2d9e1d4c6f4fcfb7f77da6594079d45ff5806bf3f89304b004

Observation 11e6ce3e-d9fb-4fa3-bc0e-51dfd75a9153 · outbound

This paper cites KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.757043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.757043Z digest=sha256:a87bcbac27c28ee0e55be0179691feb46dd8bb26512435d26b3dbcda40ff82f9

Observation 9dc6bc0e-8888-4669-a835-53b7bd753216 · outbound

This paper cites T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T00:15:39.164239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.850394Z digest=sha256:660bea8c65e6e7caabbcb918fdc4ff7ce4f94e7d243d9a05055801aff6dae419

Observation 891bda99-d29c-4b28-9199-d231d78b09fd · outbound

This paper cites Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.858364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.858364Z digest=sha256:86451f150ae2e2775904dc1ab8eab59326279e5df8a08afb81c89558a3106495

Observation 47b277f9-4a99-4dd6-8559-4ad526cc082a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.870713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.870713Z digest=sha256:ceedd4b433ab46db196a1a982c5b30a2945228cda8139b7dced718d9df4fa7f5

Observation 5ee90300-f5c7-4a68-8bda-cb214e044708 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.894887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.894887Z digest=sha256:9ce225626c3e9c866ecdcfb0fb0f6fc1daf7de03f73470132d26a2f25819b1a6

Observation d4d414c5-45ac-4f32-bad0-65de41975c9c · outbound

This paper cites Unigrpo: Unified policy optimization for reasoning-driven visual generation,.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unigrpo: Unified policy optimization for reasoning-driven visual generation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.834847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.913425Z digest=sha256:061bc0ca905bb10dde7d19e959b8c5d46d295e010ccc4c0893b4ed8e8aa470dc

Observation 83683b37-4165-45b9-8cf8-5a9dee1287ca · outbound

This paper cites Towards unified multimodal interleaved generation via group relative policy optimization, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Towards unified multimodal interleaved generation via group relative policy optimization, 2026

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:40.087957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:38.958815Z digest=sha256:5db8d313f245d2b10734408bda208f1c5099df86f5b4d1ae5ffbbeed403b66ee

Observation e96e6f37-f4f3-40cf-a77b-f52cbd65b9d6 · outbound

This paper cites WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.962629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.962629Z digest=sha256:19da30ed3092a7f11d6d13fddfc13735897a021301bc9a872b930a59bcbff67d

Observation ef9a1d55-f28e-4a1c-a270-35deafa9adec · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.966639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.966639Z digest=sha256:8b73f36aa419a3c428cefb4b8f14cbbd8647d2493c4611c7d698b438bfd4764f

Observation 025dd2f3-746e-4a69-ba7d-718778b2484f · outbound

This paper cites Bermano, and Ohad Fried.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Bermano, and Ohad Fried

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.970501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.970501Z digest=sha256:acf5390deb784005432b30a151098cf57b8561398c774c31853b51924c104b56

Observation 30ab19c4-b198-439d-8c3e-29f1c3803536 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.979572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.979572Z digest=sha256:ecd3235cbcf60dd9e496942da309053967b450814a26d9f9e408faafd0562fae

Observation aa9dca68-a0df-46d2-9ece-67f68e64cd7d · outbound

This paper cites R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.984695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.984695Z digest=sha256:e56618afe89234fc75e5297d06b257a483961bff475a88ed47b00572b5fdcf03

Observation 8691bc38-f3cb-4d90-83b1-b31ccbf23de2 · outbound

This paper cites LongCat-Image Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LongCat-Image Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.987981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.987981Z digest=sha256:c8678b0435436dca903f8e6cd5fb32816acdd9f275438b1209b126e7ef504046

Observation 1b9ca539-0aea-46de-8e61-2dd5e3472e48 · outbound

This paper cites HunyuanImage 3.0 Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation HunyuanImage 3.0 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.991808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.991808Z digest=sha256:d14f98e6d52425a59e4d54b997a0bb6b715d3a903e79b782568c942e6e09a07c

Observation a687681b-92c3-4b73-b4dd-56e7c1a35950 · outbound

This paper cites Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.574168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.009144Z digest=sha256:c6a1c75f1f55eed5f155f779c61a4397d6e6021a13c905a038c9079060a2c184

Observation b5b9e53c-ae7a-464c-b324-a1d7306013f5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3: Next-Token Prediction is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.013329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.013329Z digest=sha256:382558d577269cdfeb63ac0df1caa2e3d1742f35ec8e0c61c55f65b4073d2fdb

Observation 93e2c5a1-8cc8-4bed-ba3c-0eea7571ef07 · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.016775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.016775Z digest=sha256:9abf0afb4fb0f6458129c7c1022297a81369f0a74c73d0b7c8edeb70cb82b84a

Observation 0f0591e6-38b2-4b0a-8b80-70f60dc6d1a1 · outbound

This paper cites Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.031010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.031010Z digest=sha256:441bbf390ed4321ff63f8194cced95198051d5d8c74de58105dfbb474a8fcdbc

Observation d3b98ec3-f4b3-4feb-923e-446f55dac0ec · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ReAct: Synergizing Reasoning and Acting in Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.039044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.039044Z digest=sha256:3d1800e165b08e84c885fbea73d457c4190e8cc7d350fd03766cd565c6bd3ed1

Observation fcfbedb8-094b-49bf-b0a2-18b471803ab6 · outbound

This paper cites GenClaw: Code-Driven Agentic Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenClaw: Code-Driven Agentic Image Generation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.458384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.042271Z digest=sha256:64f6b92020beec2f54a9762bdbfb7fe1b170f58ee944f515934951d88f0d7e0d

Observation f6985950-4be7-4ee9-b73b-1a47128f7bfc · outbound

This paper cites Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.053450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.053450Z digest=sha256:f10514978e9a3ef00d7d9ca9920b9466f779065325516931566c69498ed6b963

Observation ae267b74-6094-4415-a9d5-4277b5df018e · outbound

This paper cites WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.056728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.056728Z digest=sha256:9e5c9f1b87ea44ee7492a17a7a4ee3c962798007fc1bc6bdef9d09569053818d

Observation 7aaedbb5-5bfc-4fec-8a90-e714d7e0ea61 · outbound

This paper cites Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.339882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.060537Z digest=sha256:f31784847b383e2ed57dddf82b9469d07fc931b256c3a64ea52ab5b3accdd2b6

Observation 9ee431e9-9018-4bfd-899a-05fac29ca45b · outbound

This paper cites ""You are a.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ""You are a

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.074996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.074996Z digest=sha256:dd9e8e177db8d298865933209bcca9b448f1e005ab6b182a85b7233b21089445

Observation a1b30b39-cfc0-49f3-b8f3-a70246623dcb · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:42.490196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.080550Z digest=sha256:cf447c497a08e0b803423bb8095f01f333369f89c3653edfdef253b8834deb8c

Observation 9c5b5567-0bfb-4219-8122-0e2496dba551 · outbound

This paper cites according to the document.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation according to the document

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.175268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.094934Z digest=sha256:2fc83d486e981507388ec209e40782887a4a754e75af7e146863117285b0b4eb

Observation 6fec7dc4-0cd6-4b21-9692-0984b3ba3fbf · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:41.934761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.117117Z digest=sha256:760b95f71e3ac4f0d3ed7397c7375f2f2eeb26fe01801b12702cc32c4cc078c4

Observation 6a98c200-8729-4518-b91b-59a22722cbd3 · outbound

This paper cites Avoid being overly vague.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Avoid being overly vague

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:41.657398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.121702Z digest=sha256:eb6602daf24fb395562deabfee095a775f23b1df8593fbe8017feb4fa08e3520

Observation cef43b31-d857-4142-ab78-f47d9a09ada4 · outbound

This paper cites name": "text_search.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation name": "text_search

Reference 43

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:39.243712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:15:39.125513Z digest=sha256:9a17d091a9f1eba98ec3d51231890b49337ec86b2154f3ab015d3e686af3bb99

Observation 07d4a842-96f5-4ece-890f-2afc96a33d97 · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.947081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.947081Z digest=sha256:e7515b51da7cbd1fcadb9f5ee5f061663434a3f9573b999927fe7e19ba7e2f89

Pith citing papers

No inbound Pith citation observations are available.