Pith. sign in

Paper Citation Record · LEDGER

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation

As of 9 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2608.04436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04436 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:15:39.125513Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4cced5a-a747-4f11-ab60-922e101f0195 · outbound

This paper cites Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:15:41.386031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:37.003251Z digest=sha256:a7f58fc97c736131b5eb889ad7fb11c80a0dd717762ef066860e83256ff9b2b8

Observation 589322b1-f649-457a-802b-72c2b985e498 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.175322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.175322Z digest=sha256:8f5562b678a315a05c0217c77ab53f2f7b1fdd2163c090658b002d9878e37761

Observation 304b4280-e528-4849-a48b-0aab005f9f19 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.290999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.290999Z digest=sha256:e050f56e6666912788178d3615b7a2e63798de23cef1fcabe97d0b95128f6246

Observation bd365d23-98d6-4c9b-b397-2e92feba0156 · outbound

This paper cites Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unify-agent: A unified multimodal agent for world-grounded image synthesis, 2026

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.407731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.407731Z digest=sha256:555037ed7bc4c8084258accf2975b82fe1daea9425f3e0e6b8cb395b8270ead4

Observation 653e64f5-e909-4828-89fa-a148fcc84429 · outbound

This paper cites GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenEvolve: Self-Evolving Image Generation Agents via Tool-Orchestrated Visual Experience Distillation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.549104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.549104Z digest=sha256:3cca40cd2b4da012386ba64420057a1f80a002958e6fbd29f54d81ab889f30c9

Observation 0cbe4b8d-b2e4-4b47-8431-493a4cf4bce4 · outbound

This paper cites Re-Imagen: Retrieval-Augmented Text-to-Image Generator.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Re-Imagen: Retrieval-Augmented Text-to-Image Generator

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.706082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.706082Z digest=sha256:e7671c5076bc00121b547489fecf7a788004af0506e74e8a7bbf6083297d8477

Observation 045cae11-6dd4-45f1-ab2b-70bb179ab2f9 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.847345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.847345Z digest=sha256:f8c170612a78c119a67fdd3d01f2461eedf850912d4946972032c19e56dba565

Observation 3ae2b93c-2391-4649-8aea-bf48113677a2 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3.5: Native Multimodal Models are World Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:37.958795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:37.958795Z digest=sha256:1f018650e9cbc02baa478943b361d2fcefa4524abe709ba19f4c24d98c0dab98

Observation 3081dbd7-c498-4520-b610-47e6b2990731 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.098358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.098358Z digest=sha256:f3c79147dff97c3fb36e2207664c2905a80932371bb2b225c287d899f8569986

Observation 97a849ac-d1d9-43d2-beea-addf51bf98d6 · outbound

This paper cites Gen-Searcher: Reinforcing Agentic Search for Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Gen-Searcher: Reinforcing Agentic Search for Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.217313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.217313Z digest=sha256:fa75882befdd2ddc1d79f8bc503e1fd55489d791ee499a5157c8f75b49ed592c

Observation b8e6eff9-412a-43a9-9940-376db9b11c24 · outbound

This paper cites Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Commonsense-T2I Challenge: Can Text-to-Image Generation Models Understand Commonsense?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.364743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.364743Z digest=sha256:8aab5bb495f6f3f9fcac5a523a473bdb91a662b362c21f7925eb1ed22a8a9ddd

Observation 5da3c431-64e4-40f9-8d53-0f79d2030517 · outbound

This paper cites Demystifying Flux Architecture.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Demystifying Flux Architecture

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.504110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.504110Z digest=sha256:86b5dc9ae2b31bed78768d20985d742a4a3aff814c576ad725b76ea0daefd411

Observation fe808141-0b6a-4c12-9b3f-0087ebacd1b8 · outbound

This paper cites Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Beyond words and pixels: A benchmark for implicit world knowledge reasoning in generative models, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.622986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.622986Z digest=sha256:c7558774f87f1801110d8602e51a7d322fad4a6c9c54d22bad4b17f1aeda9795

Observation 11e6ce3e-d9fb-4fa3-bc0e-51dfd75a9153 · outbound

This paper cites KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation KITTEN: A Knowledge-Intensive Evaluation of Image Generation on Visual Entities

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.757043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.757043Z digest=sha256:0f9dc4bc75d7bbafd0b38c47c34981d15b8a2f850013cbf6272171c8924a21fd

Observation 9dc6bc0e-8888-4669-a835-53b7bd753216 · outbound

This paper cites T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation T2i-factualbench: Benchmarking the factuality of text-to-image models with knowledge-intensive concepts

Reference 15

Resolution
verified exact
doi, observed 2026-08-07T00:15:39.164239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:38.850394Z digest=sha256:5d5ddd0c87a6ecce66de9d7fd5d2bab40da0e5a44367d45e2316d79956d01b57

Observation 891bda99-d29c-4b28-9199-d231d78b09fd · outbound

This paper cites Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genagent: Scaling text-to-image generation via agentic multimodal reasoning, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.858364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.858364Z digest=sha256:74234f9cc753f77b92dd27cf6f8ee8de98365b403b796740eafbdb891df8a2ef

Observation 47b277f9-4a99-4dd6-8559-4ad526cc082a · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.870713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.870713Z digest=sha256:04e7a01f481d9f773fd77e08b81e5485e0688fd7026161e06edcbd93f024030e

Observation 5ee90300-f5c7-4a68-8bda-cb214e044708 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.894887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.894887Z digest=sha256:d882a7dd53c7fc607d43013ffc91e70bb6ca81cd85a0debaed95c6732783dc98

Observation d4d414c5-45ac-4f32-bad0-65de41975c9c · outbound

This paper cites Unigrpo: Unified policy optimization for reasoning-driven visual generation,.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unigrpo: Unified policy optimization for reasoning-driven visual generation,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.834847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:38.913425Z digest=sha256:49ccb5f8fca63dc62e6222b05d954b41f60ac1c6c07dcee41bf44f74103f1839

Observation 83683b37-4165-45b9-8cf8-5a9dee1287ca · outbound

This paper cites Towards unified multimodal interleaved generation via group relative policy optimization, 2026.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Towards unified multimodal interleaved generation via group relative policy optimization, 2026

Reference 20

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:40.087957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:38.958815Z digest=sha256:a2d7ea10ecf9f81141bada14f98d32c124cf8583da251c6920761e1ac03ad7bc

Observation e96e6f37-f4f3-40cf-a77b-f52cbd65b9d6 · outbound

This paper cites WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.962629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.962629Z digest=sha256:3d970346a4dc8971080b5e3b7aedad64cb0d21b1efc037ad3aa4960808a4f37e

Observation ef9a1d55-f28e-4a1c-a270-35deafa9adec · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.966639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.966639Z digest=sha256:01d13406b5c682a84966436e41410d749487786df774274b0afdeefcc1e28aa5

Observation 025dd2f3-746e-4a69-ba7d-718778b2484f · outbound

This paper cites Bermano, and Ohad Fried.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Bermano, and Ohad Fried

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.970501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.970501Z digest=sha256:4cd402d15abeb691655679127c2d3f97435739930660ab4be735b4a554847599

Observation 30ab19c4-b198-439d-8c3e-29f1c3803536 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.979572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.979572Z digest=sha256:7cd41578ade894e4c78568d8de8eb90968071a887f681f1dc723be200a900ad2

Observation aa9dca68-a0df-46d2-9ece-67f68e64cd7d · outbound

This paper cites R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.984695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.984695Z digest=sha256:fbb0d5040c624584528b4e7ab5201b8a93d3bef48776e09e0db737a5b74ac089

Observation 8691bc38-f3cb-4d90-83b1-b31ccbf23de2 · outbound

This paper cites LongCat-Image Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation LongCat-Image Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.987981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.987981Z digest=sha256:8b7e77d511db5dec2900343b653379caa71141cc2b12e3ede7de46f8e9b091ec

Observation 1b9ca539-0aea-46de-8e61-2dd5e3472e48 · outbound

This paper cites HunyuanImage 3.0 Technical Report.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation HunyuanImage 3.0 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.991808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.991808Z digest=sha256:d0060cee626672b2e883f4d78c6fb8653242c9b70d9bfe3b8107109b383d51de

Observation a687681b-92c3-4b73-b4dd-56e7c1a35950 · outbound

This paper cites Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Search Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.574168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.009144Z digest=sha256:73828f4335ecbe3219dbb2632b10f6c2fe30725f8868f68736a484316ba554d0

Observation b5b9e53c-ae7a-464c-b324-a1d7306013f5 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Emu3: Next-Token Prediction is All You Need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.013329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.013329Z digest=sha256:ab17be2b305bf26e06570e5d14df394b8d57502085e10dfe362db78434a894bb

Observation 93e2c5a1-8cc8-4bed-ba3c-0eea7571ef07 · outbound

This paper cites Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.016775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.016775Z digest=sha256:64a7a618af17546567528e336eb81af86098e6777df10f20f9c0207654a70fbe

Observation 0f0591e6-38b2-4b0a-8b80-70f60dc6d1a1 · outbound

This paper cites Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.031010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.031010Z digest=sha256:c6454792a46f88934f3000ad71f45164036b60c63ba7ccdd97a73aaf781962af

Observation d3b98ec3-f4b3-4feb-923e-446f55dac0ec · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ReAct: Synergizing Reasoning and Acting in Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.039044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.039044Z digest=sha256:69acc525d0a4ab08517e0477f2a1dbb7eed76e22578271d958c7ea06680f55be

Observation fcfbedb8-094b-49bf-b0a2-18b471803ab6 · outbound

This paper cites GenClaw: Code-Driven Agentic Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation GenClaw: Code-Driven Agentic Image Generation

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.458384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.042271Z digest=sha256:9fb71482bac4e6c7ccab14ef4f7adfd30b54d4ca411bee59ba512cc37809a280

Observation f6985950-4be7-4ee9-b73b-1a47128f7bfc · outbound

This paper cites Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Genpilot: A multi-agent system for test-time prompt optimization in image generation, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.053450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.053450Z digest=sha256:dfbdabfc97d5306ef12bc7e0167c91c4e1445cefd0fd7d8726337055e4b0a4fb

Observation ae267b74-6094-4415-a9d5-4277b5df018e · outbound

This paper cites WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation WorldGenBench: A World-Knowledge-Integrated Benchmark for Reasoning-Driven Text-to-Image Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.056728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.056728Z digest=sha256:d9ff7cf073d02b70d8eec1a7efa872588231fe5201ca24e26a0f623cdb3400df

Observation 7aaedbb5-5bfc-4fec-8a90-e714d7e0ea61 · outbound

This paper cites Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:15:39.339882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.060537Z digest=sha256:ca8ab54c2e0adcb48c306581452e55c9feaf62a3ffb0f027f13b004f118bdff0

Observation 9ee431e9-9018-4bfd-899a-05fac29ca45b · outbound

This paper cites ""You are a.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation ""You are a

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:39.074996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:39.074996Z digest=sha256:e133c5e262cf28dbaad9cc1d5a029fa97bd6393b34161b55afba13f83fb76867

Observation a1b30b39-cfc0-49f3-b8f3-a70246623dcb · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:42.490196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.080550Z digest=sha256:4a51210282463a39408c79f3d8460497e31f6a7965fdddaa82c649d80d80eb16

Observation 9c5b5567-0bfb-4219-8122-0e2496dba551 · outbound

This paper cites according to the document.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation according to the document

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:42.175268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.094934Z digest=sha256:f6873441deb63c556183d0d89a03825d6b0f8e516e7810b435c00d15d1f1ea51

Observation 6fec7dc4-0cd6-4b21-9692-0984b3ba3fbf · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:15:41.934761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.117117Z digest=sha256:c3f579d36b1fa47d3408365a7468d561c7e5bf261b1f69c99814528cd6174253

Observation 6a98c200-8729-4518-b91b-59a22722cbd3 · outbound

This paper cites Avoid being overly vague.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Avoid being overly vague

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:15:41.657398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.121702Z digest=sha256:e11f9c6a777ed9b6c64d5fc90747a9fef872aac7c9f863f2325187be1604a4c7

Observation cef43b31-d857-4142-ab78-f47d9a09ada4 · outbound

This paper cites name": "text_search.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation name": "text_search

Reference 43

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:15:39.243712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T00:15:39.125513Z digest=sha256:bc5d52b30a7628d797cd846ccbb11de426642018f5f86dd30e2a49237899d5b5

Observation 07d4a842-96f5-4ece-890f-2afc96a33d97 · outbound

This paper cites an unresolved cited work.

ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Unresolved cited work

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:38.947081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:38.947081Z digest=sha256:edd2a06c840e108e7ed6aeee580c4143d5ef6a99b74568c548b6d67b9ac3cc56

Pith citing papers

No inbound Pith citation observations are available.