Pith. sign in

Paper Citation Record · LEDGER

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

As of 7 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2608.04436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04436 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:15:39.125513Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4cced5a-a747-4f11-ab60-922e101f0195 · outbound

This paper cites Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:15:41.386031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:37.003251Z digest=sha256:dd6f66e074015eebac98e11a17209eeebc5a457e6cd06ab1dedd7603d6c30f86

Observation 589322b1-f649-457a-802b-72c2b985e498 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.175322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.175322Z digest=sha256:2cb9aaccd718cf239ab43fbdbf0932bcb58ef326a9619d76149c54424b4679cf

Observation 304b4280-e528-4849-a48b-0aab005f9f19 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.290999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.290999Z digest=sha256:84747274dde963e3a7260c074fc909adda43870aba6a5554f8c65199f0cc2532

Observation bd365d23-98d6-4c9b-b397-2e92feba0156 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.407731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.407731Z digest=sha256:3d385051013c0e1c476bd1bbbd9599cdbcb031ab7f1251742b84ec7c4ff3a8c7

Observation 653e64f5-e909-4828-89fa-a148fcc84429 · outbound

This paper cites GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.549104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.549104Z digest=sha256:d2af4fe59ad67c6d917ca585b6cc77e024e412db0c3af6278ebee4e0f45ad959

Observation 0cbe4b8d-b2e4-4b47-8431-493a4cf4bce4 · outbound

This paper cites Re-Imagen: Retrieval-Augmented Text-to-Image Generator.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Re-Imagen: Retrieval-Augmented Text-to-Image Generator

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.706082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.706082Z digest=sha256:d72b9e92535e64a71ecd6f511c5e299de537b24055929301617c1820da542251

Observation 045cae11-6dd4-45f1-ab2b-70bb179ab2f9 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.847345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.847345Z digest=sha256:9442b2c1bc157eb060dd3fb2a1cb072b1bf582f8499210f5beeb5143fcaeccc5

Observation 3ae2b93c-2391-4649-8aea-bf48113677a2 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3.5: Native Multimodal Models are World Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.958795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.958795Z digest=sha256:d81fd406a7b460cfe17ab2d5119ad1654ac345412739fa1488e5da3b80f4cc21

Observation 3081dbd7-c498-4520-b610-47e6b2990731 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.098358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.098358Z digest=sha256:76003e76e82bd901457b3117e8c539b3a0802c14c102a15e6eddc15ed88578cd

Observation 97a849ac-d1d9-43d2-beea-addf51bf98d6 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.217313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.217313Z digest=sha256:31bdf15d76f47a95ef3463f14554a8de9dd88f8c3205d34c3d4ac847090b918d

Observation b8e6eff9-412a-43a9-9940-376db9b11c24 · outbound

This paper cites Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.364743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.364743Z digest=sha256:46c2b59183538f62d98c3d4c77ef3b3ab958bacfea293321673f92946bdf3465

Observation 5da3c431-64e4-40f9-8d53-0f79d2030517 · outbound

This paper cites Demystifying Flux Architecture.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Demystifying Flux Architecture

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.504110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.504110Z digest=sha256:ab3ea3a505a7d07d8d7b44da0557beb85fe3c48c243bdb40beafb47eb38c17ef

Observation fe808141-0b6a-4c12-9b3f-0087ebacd1b8 · outbound

This paper cites Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.622986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.622986Z digest=sha256:5c0190a43c822d9be5166452b0797721446de70873b7fc186be94eb8f064e474

Observation 11e6ce3e-d9fb-4fa3-bc0e-51dfd75a9153 · outbound

This paper cites KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.757043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.757043Z digest=sha256:09a77b1506dac6c94fa3fb2139f3e6f808b4ca07bed842c6f6c2434e24540475

Observation 9dc6bc0e-8888-4669-a835-53b7bd753216 · outbound

This paper cites T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T00:15:39.164239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:38.850394Z digest=sha256:2dc835f2f2c60daf9fd4f30035aec6602cec5d5d0e2e00a9280de390707dea33

Observation 891bda99-d29c-4b28-9199-d231d78b09fd · outbound

This paper cites Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.858364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.858364Z digest=sha256:48a302dcdef90ad5e657c9c7e3d74679d1d076fcb480f5433bf1d2c4047c4348

Observation 47b277f9-4a99-4dd6-8559-4ad526cc082a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.870713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.870713Z digest=sha256:030e50be749b2e46f174e43bcd1fb4b131f8ffb6738f1a05b6f60bc512048949

Observation 5ee90300-f5c7-4a68-8bda-cb214e044708 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.894887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.894887Z digest=sha256:c8a202fb9b45b9dbf07b1edc111a552775868a3f69db46d0a11cc24f2bfd77b9

Observation d4d414c5-45ac-4f32-bad0-65de41975c9c · outbound

This paper cites Unigrpo: Unified policy optimization for reasoning-driven visual generation,.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unigrpo: Unified policy optimization for reasoning-driven visual generation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.834847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:38.913425Z digest=sha256:bc4188e79a473adb62b1f1be4aabc17709987c6add9db22061d1606cc072ce01

Observation 83683b37-4165-45b9-8cf8-5a9dee1287ca · outbound

This paper cites Towards unified multimodal interleaved generation via group relative policy optimization, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Towards unified multimodal interleaved generation via group relative policy optimization, 2026

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:40.087957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:38.958815Z digest=sha256:97c5ca96abb95dacb9fffc0be9df00d3c8dfa14d076123186279e88f84ea7b45

Observation e96e6f37-f4f3-40cf-a77b-f52cbd65b9d6 · outbound

This paper cites WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.962629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.962629Z digest=sha256:097f35ca700ee8a3be2ad116b5d9ad6c7b786126a11a50b6c7702e54dc50e1b1

Observation ef9a1d55-f28e-4a1c-a270-35deafa9adec · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.966639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.966639Z digest=sha256:8d94febabf4e00dbebaeb9ddc649a24c64f34f93b4b253721775ac1d6f4e5dcb

Observation 025dd2f3-746e-4a69-ba7d-718778b2484f · outbound

This paper cites Bermano, and Ohad Fried.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Bermano, and Ohad Fried

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.970501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.970501Z digest=sha256:06dc28a66e1c01b062138c8b2a004de3e453e5523dc3f0df99dddd03c25c37da

Observation 30ab19c4-b198-439d-8c3e-29f1c3803536 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.979572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.979572Z digest=sha256:1c524a45ed9919a0c7c356c275c7ffa1bf7cbf17994a3719358d1a927457f806

Observation aa9dca68-a0df-46d2-9ece-67f68e64cd7d · outbound

This paper cites R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.984695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.984695Z digest=sha256:7d5d5fee97e9fcf66c797776f1b2ac7df7918fc16cdee902ae6bf7db531fa58a

Observation 8691bc38-f3cb-4d90-83b1-b31ccbf23de2 · outbound

This paper cites LongCat-Image Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LongCat-Image Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.987981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.987981Z digest=sha256:51535881f7a7ba3476c67426a1edb85d15c995d7fcc0453a6079257bfcb0b817

Observation 1b9ca539-0aea-46de-8e61-2dd5e3472e48 · outbound

This paper cites HunyuanImage 3.0 Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation HunyuanImage 3.0 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.991808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.991808Z digest=sha256:d0b13a65d4d9a308e38319dc5cd6390794a4006297f69cbf608abab7f0c1b5a0

Observation a687681b-92c3-4b73-b4dd-56e7c1a35950 · outbound

This paper cites Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.574168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.009144Z digest=sha256:758e32a02dfee88ef9e7f7bf88c966c14b170d136b5b4bf63216f070f57171dd

Observation b5b9e53c-ae7a-464c-b324-a1d7306013f5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3: Next-Token Prediction is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.013329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.013329Z digest=sha256:c28928b56b41bf1a3923590ca5cedc953e21423053f91d7a4cb3a405728cc9cd

Observation 93e2c5a1-8cc8-4bed-ba3c-0eea7571ef07 · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.016775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.016775Z digest=sha256:87ba53361b7db6b911543daa96c351bec0c2f8c5c15d33a1b59645c41f5ec695

Observation 0f0591e6-38b2-4b0a-8b80-70f60dc6d1a1 · outbound

This paper cites Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.031010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.031010Z digest=sha256:ff8615e9a938da288c68a3d22d2513c274aabb984ef7332f6df34eeaee5e36cd

Observation d3b98ec3-f4b3-4feb-923e-446f55dac0ec · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ReAct: Synergizing Reasoning and Acting in Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.039044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.039044Z digest=sha256:34533217265c0cb8d4ef55c37d61ae19ef66c6298dfe00d730c8938517716184

Observation fcfbedb8-094b-49bf-b0a2-18b471803ab6 · outbound

This paper cites GenClaw: Code-Driven Agentic Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenClaw: Code-Driven Agentic Image Generation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.458384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.042271Z digest=sha256:07d123df73d01152a8b7301bdf10f4a89c58cc7666f950390858f34bd266d14d

Observation f6985950-4be7-4ee9-b73b-1a47128f7bfc · outbound

This paper cites Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.053450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.053450Z digest=sha256:a5061e4b6d4130077bfd085c469fa04e3e948dfc1daa9f729452f57968173e2e

Observation ae267b74-6094-4415-a9d5-4277b5df018e · outbound

This paper cites WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.056728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.056728Z digest=sha256:661106e4c952f6b6718ba42e7b4a12cbdffb8ee96737f817e428133b2349e38b

Observation 7aaedbb5-5bfc-4fec-8a90-e714d7e0ea61 · outbound

This paper cites Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.339882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.060537Z digest=sha256:bcef0d8f24d07c2ace671173c507458c62afd1ee3f92c3ad15aceb397f65fb11

Observation 9ee431e9-9018-4bfd-899a-05fac29ca45b · outbound

This paper cites ""You are a.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ""You are a

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.074996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.074996Z digest=sha256:e27e47e752778cf8390905e771fb15d43f5719046e93e2b379129e1af6b57f49

Observation a1b30b39-cfc0-49f3-b8f3-a70246623dcb · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:42.490196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.080550Z digest=sha256:01e69111060b51e2ae32c8001102ebc531d773ce33cb273db030e0aef16e9471

Observation 9c5b5567-0bfb-4219-8122-0e2496dba551 · outbound

This paper cites according to the document.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation according to the document

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.175268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.094934Z digest=sha256:c83fc23bc82fc332803326fc00b7286ed567f86e1db92a79e61e6c0b3d64b618

Observation 6fec7dc4-0cd6-4b21-9692-0984b3ba3fbf · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:41.934761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.117117Z digest=sha256:66135841f65d4d2336803b603a071e8c3f3c0f943e15a16a648de42bd2610e91

Observation 6a98c200-8729-4518-b91b-59a22722cbd3 · outbound

This paper cites Avoid being overly vague.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Avoid being overly vague

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:41.657398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.121702Z digest=sha256:cc343e632243c4b33e9114c14d83f1ea7f6a6e0246040c819cd2080cfb138b60

Observation cef43b31-d857-4142-ab78-f47d9a09ada4 · outbound

This paper cites name": "text_search.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation name": "text_search

Reference 43

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:39.243712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:15:39.125513Z digest=sha256:1df94f4ff52e5f980abbd0413c576aad087259a577b1b92494a2d8516cd8baab

Observation 07d4a842-96f5-4ece-890f-2afc96a33d97 · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.947081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.947081Z digest=sha256:3c978c726af2b275a5309592ad27c556c8a2eed83c52d1b5f88722a00b261ed5

Pith citing papers

No inbound Pith citation observations are available.