Pith. sign in

Paper Citation Record · LEDGER

TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 17 inbound Pith citation observations for arXiv:2506.02161.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02161 v3

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:33:20.868273Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:44:16.332457Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:39:58.183804Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7ff905bf-341b-45f3-9578-3604d525bb03 · outbound

This paper cites High-Resolution Image Synthesis With Latent Diffusion Models.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? High-Resolution Image Synthesis With Latent Diffusion Models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.569799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.699215Z digest=sha256:23928c4acf7eecd577dec26991f4c315c90a6ef7e7e9882fc88003fd327d8de7

Observation 8e7e68a8-6871-466e-9dff-6bbe2f888771 · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.558683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.703880Z digest=sha256:2ce8cb05ffe594243b2e18d972a1274a9f5888b8b1ba355d037f10e6df1f6a0c

Observation 13edd74b-bb06-4021-be70-92906ebea776 · outbound

This paper cites PixArt- Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation, March 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PixArt- Σ: Weak-to-Strong Training of Diffusion Transformer for 4K Text-to-Image Generation, March 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.546906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.707732Z digest=sha256:18d77684070b59093a4a8e69f42689a2908280bbda8cad28e51c79d75301f286

Observation d76aa8f4-7e3b-49d9-a893-3801d9858507 · outbound

This paper cites PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models, January 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PIXART-δ: Fast and Controllable Image Generation with Latent Consistency Models, January 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.535901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.711497Z digest=sha256:8309a7ee57d62d2bc93fe3594c668af81ec47d5e92ba72f789a2503109cd40d7

Observation ce296ab9-d4b2-4730-8159-a8ef43c3afe0 · outbound

This paper cites PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis, December 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? PixArt-$α$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis, December 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.524832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.715348Z digest=sha256:2d093f0d2849abf6e55c183a954aed168daff805d2b6e50636d469ed6a52f687

Observation 5d428585-9253-442d-a054-c3ffd40912b7 · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation, February 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation, February 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.512514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.719099Z digest=sha256:6265eded1ca18d30c6925e952f175feb067f2989f116cf1e538f0611f3900172

Observation aa994f09-1958-4c7a-b53f-16ca9548f966 · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models, October 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.500419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.723623Z digest=sha256:681ee46587b6c25f730f69f2a2c2676b7db6301080c6fc4a5441cdbae5e0d61f

Observation c1da543c-d3b7-4b5a-ac5b-ceefcc5f0035 · outbound

This paper cites Flux.https://github.com/black-forest-labs/flux, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Flux.https://github.com/black-forest-labs/flux, 2024

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.727512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.727512Z digest=sha256:c67efd7f94e2189efab99e8a36d7dfa6d5878e7fe1886b322ec4af073d8e3885

Observation fb5ae75c-c0d9-4796-aff1-4b1ff5056002 · outbound

This paper cites SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation, March 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.482294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.731366Z digest=sha256:cd5dc2ca00bc9548af14a33492d4d960b2d0c57cf0a079114c1bb9cb6f55cba4

Observation 883a058c-3e0b-40cf-8789-dfc2789a4da8 · outbound

This paper cites SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer, March 2025

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.471370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.734808Z digest=sha256:5017b7bef20bd06da7dad71e171c5788a63f6c1a2f3545278f9efe3cf68f3aea

Observation 8133f7a2-0acf-4c60-a818-418394433491 · outbound

This paper cites SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers, October 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.459949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.738287Z digest=sha256:7c8def2b7abf6ab98244d7150e431cc4385e8457186b749b282dfb7ac8274010

Observation d9199a8c-be61-4d0e-afb5-47c79af3ffe6 · outbound

This paper cites Improving Image Generation with Better Captions.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Improving Image Generation with Better Captions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.448476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.742158Z digest=sha256:af86966f037fef41d02dd77403014987ca7680ffa819bd51f7e65b72f5bb9c43

Observation 638667a5-baac-41ff-b36a-d16d6f44466b · outbound

This paper cites Imagen 3, December 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagen 3, December 2024

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.436139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.745747Z digest=sha256:77d547ab25ef9cf20e4bcc0fc801fabde5f4c2107a3e7ced79d782c17187cfab

Observation 44747427-b8aa-4b85-847b-e793814bd62c · outbound

This paper cites Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT, June 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT, June 2024

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.423834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.749257Z digest=sha256:f43ccdb00ec798ec967b2b7e9c6694a6275ad9c6ed985e669c22f2c88bb6fc70

Observation 9698db62-1f56-4938-ac68-436fc63f0d99 · outbound

This paper cites Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding, May 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding, May 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.412266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.752517Z digest=sha256:3753dfd8cb9cf49d3efa968cc2fef159f38655c72faa6b4a2898e0b592cc2cd3

Observation 28305e17-2cef-47c9-af23-42337bbf3015 · outbound

This paper cites Autore- gressive Model Beats Diffusion: Llama for Scalable Image Generation, June 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Autore- gressive Model Beats Diffusion: Llama for Scalable Image Generation, June 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.400614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.756670Z digest=sha256:ab19bf54e2a336634b2cb4f1c66049f2922fb9cd7089eded8946891a250a579e

Observation 3c2a6e9d-65f4-439a-ae43-a6733840b6cc · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.388705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.760152Z digest=sha256:99378d167daa8149d53e2c5635a2b511a78f99d4ea90cd999d11d293c2ee45ec

Observation 4ae405c8-f22f-44b4-b98d-85b2d13d0a1c · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation, October 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.377520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.763663Z digest=sha256:8f722e73b78d86417b04e2cd8283ad570cd39a463156a7cfaf87909647463460

Observation d1e3c665-6d92-4bed-8ed5-a72308b59dd8 · outbound

This paper cites Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.366781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.767179Z digest=sha256:3a5e8664cdb0fec88c93cbf7b378051303babbe4fd797899e7d7e30d25a467a1

Observation 27e0a2a4-a500-46c6-bd41-3bc3a3af796c · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.355174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.770439Z digest=sha256:95cf3c3f63e0b715881e948c73e03d04f46b1fcc74605b59169c8b355900ce10

Observation 83c040ef-08e5-41b0-af2d-d48571acb7df · outbound

This paper cites Lumina-mgpt: Illuminate flexible photorealistic text-to-image generation with multimodal generative pretraining, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lumina-mgpt: Illuminate flexible photorealistic text-to-image generation with multimodal generative pretraining, 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.344009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.773908Z digest=sha256:bf545277baa58a184755ca034e501422079ff56be6abfb422072a18428cdb3ad

Observation d51ba9b8-8998-4242-9582-26b28514a842 · outbound

This paper cites T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.777967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.777967Z digest=sha256:8fcc7c56a9ae46d6873aef7de4f454e1b7d547516e31e8a58447d41527db777a

Observation 553d2b69-b567-4d94-9e80-6517d6fa6cf0 · outbound

This paper cites Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.782033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.782033Z digest=sha256:e8bbae5460c16d4b6ddb0c6ee6258b17ae9bf96542c2dc197e74573965613efa

Observation 5740b8a0-7f7d-4d14-bbd8-a9ff0c6050f9 · outbound

This paper cites Delving into rl for image generation with cot: A study on dpo vs.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Delving into rl for image generation with cot: A study on dpo vs

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.332109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.786270Z digest=sha256:1c66028f5ffcc0f67d5263b9da40f4fad09964c9c1037f49eefdcde8e9fd533d

Observation 29c3bb24-c953-4a06-a08b-71225aea3583 · outbound

This paper cites Mavis: Mathematical visual instruction tuning with an automatic data engine, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Mavis: Mathematical visual instruction tuning with an automatic data engine, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.789700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.789700Z digest=sha256:2ac4be897f569d42c48367154d388a5a39ac53b01f2d3589a3e035216dfc763c

Observation 364380b4-c6ac-44bd-b570-7102361c36ee · outbound

This paper cites GPT-4o System Card.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.793233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.793233Z digest=sha256:64b8f0076eff361d2d055c08f12fd706f255431f04239a8099dcb7f4f5e33c91

Observation 4d8027c9-2d2e-49c1-95c2-c923e1d5c232 · outbound

This paper cites Imagen 3.arXiv preprint arXiv:2408.07009, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagen 3.arXiv preprint arXiv:2408.07009, 2024

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.796866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.796866Z digest=sha256:eabc33226dd7d4ece59728118cc071962bdb5d80408233b4d28b792ce363c2dc

Observation 38ea550e-9121-49d3-8525-8b8566278a0b · outbound

This paper cites Midjourney.https://www.midjourney.com/, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Midjourney.https://www.midjourney.com/, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.312181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.800613Z digest=sha256:83d5103b1484d76323c14e53b66fd21ad8a1a49935873fd70b74906d28e9a873

Observation 80ac617c-712c-41de-80d3-f01aff1c76fc · outbound

This paper cites Evaluating text-to-visual generation with image-to-text generation, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Evaluating text-to-visual generation with image-to-text generation, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.804182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.804182Z digest=sha256:fbde725af5dfffb734eeaaa234a21c3d38d372cb8ef0bde5d5f8f083054cba3e

Observation 86971d00-3aaf-4745-8f7c-19e4f6ee8703 · outbound

This paper cites Human preference score v2: A solid benchmark for evaluating human preferences of text-to-image synthesis, 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Human preference score v2: A solid benchmark for evaluating human preferences of text-to-image synthesis, 2023

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.291926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.807563Z digest=sha256:e23bae8ef31cae941e1e4061dff9899f6ca3e36cb51185776d7b1047c7d61241

Observation 179e58ca-b814-4fec-b29f-d1697ca4881e · outbound

This paper cites Visionreward: Fine-grained multi-dimensional human preference learning for image and video generation, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Visionreward: Fine-grained multi-dimensional human preference learning for image and video generation, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.280082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.811059Z digest=sha256:33171e7910ef5e305ec993c4a0282cb6b94249b33fe93cd146ab6ed7820b8faa

Observation 0b779615-ad05-463c-9eeb-48d67760f83d · outbound

This paper cites T2i-compbench++: An enhanced and comprehensive benchmark for compositional text-to-image generation, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? T2i-compbench++: An enhanced and comprehensive benchmark for compositional text-to-image generation, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.267638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.814412Z digest=sha256:afedf9c88c14227f30a44f4ede16fa3e283f62df6f2027db66a2ca54202a0383

Observation 335d14f5-8436-407b-9f55-189d1330a751 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment, 2023.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Geneval: An object-focused framework for evaluating text-to-image alignment, 2023

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.817874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.817874Z digest=sha256:ca96d3ad57517dfd79a546e5a040cf5302b8f34a0513acf6dedc9899d35013b1

Observation 43520612-cb4f-4e8c-afe1-d38d1ebc2401 · outbound

This paper cites Genai-bench: Evaluating and improving compositional text-to-visual generation, 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Genai-bench: Evaluating and improving compositional text-to-visual generation, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.246996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.821342Z digest=sha256:b26458906f8d6e6054c07330dcf6a46a9610d3051ae79a771bf0cb9d7a2c0b6b

Observation 92376c32-a6e6-4ef3-8acd-2cd6bdb6b3e2 · outbound

This paper cites Improving image generation with better captions.Computer Science.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Improving image generation with better captions.Computer Science

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:20.824806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:20.824806Z digest=sha256:050de2868e7d70a417a2add8a569be593afa2b601d5b8232b94ea23e975eddf8

Observation d2a9b0fa-ba67-4406-a5f5-add6511584e3 · outbound

This paper cites LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization, March 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? LightGen: Efficient Image Generation through Knowledge Distillation and Direct Preference Optimization, March 2025

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.227046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.828532Z digest=sha256:c349c0fb1f61e307b44f5d93e1222cc752b3d127a9f1017449fcefc834c33379

Observation 21cad2aa-f441-4178-949a-b0481e52e804 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens, October 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.214959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.831847Z digest=sha256:e3e3035b5edd95746206994b234beaa1eb8a46561e394b6183bcc3fdd16d9296

Observation 746af4be-104a-40f3-960b-b608a9d7629a · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning, March 2022.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? CLIPScore: A Reference-free Evaluation Metric for Image Captioning, March 2022

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.202989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.835257Z digest=sha256:577ea0c7929a72aa319089e90c355813e7caddcd696431f04ede26acf94e1fc9

Observation fd884cf7-ad9b-42c9-9dfd-821e887dd63c · outbound

This paper cites Imagine-e: Image generation intelligence evaluation of state-of-the-art text-to-image models, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Imagine-e: Image generation intelligence evaluation of state-of-the-art text-to-image models, 2025

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.190697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.838627Z digest=sha256:5fe5fd74003f8e9019b9ef26304a15bee88a578c21c68d93d291b1c10d89f06f

Observation 3a90cc84-1aa0-4993-9884-a014d9ecaccd · outbound

This paper cites Lex-art: Rethinking text generation via scalable high-quality data synthesis, 2025.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Lex-art: Rethinking text generation via scalable high-quality data synthesis, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.179154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.842182Z digest=sha256:70b6a8cbf46447465b6105e7966d44dcb63105b68f2ddf672dccf5b520a2b586

Observation 0b22b191-896d-4a76-9abd-cc14c3c140d2 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation, October 2024.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Show-o: One Single Transformer to Unify Multimodal Understanding and Generation, October 2024

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.166941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.845592Z digest=sha256:585bc7fe384283834e06855e34296d94bea61625888d8220387a67a28f0abf4d

Observation 40e9853e-f855-4e1c-a50c-1adb1e4d4582 · outbound

This paper cites yes" or.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? yes" or

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:33:21.153739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.849324Z digest=sha256:80de3f7bb7a015b5f165e6d033357d7341c006e026d0d32e34a1f08967c716ef

Observation 91e5dbe3-bc54-4686-8356-dcf7b0b11bc2 · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.142023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.853261Z digest=sha256:7550ee11d1564a440dcf782d957a2cf825c7bf1870ab7e606aaf6912bfc87a83

Observation 79d57472-ea30-4575-9ae7-717b21dc83ba · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.131023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.857139Z digest=sha256:87b2c84746ef83fe3cf4191a0204f0e245b42ab1c91417a75d9461745c29cc6e

Observation e0e8501f-69f6-40cf-8445-7ac10d0a5034 · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.119843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.860958Z digest=sha256:9515649cda78bb5bb7066cc19868ba68f7158bb59306fa48df1b79d3424be9c3

Observation ceb6d571-4669-494b-b743-5be1ca43a0ef · outbound

This paper cites an unresolved cited work.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:33:21.107473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.864701Z digest=sha256:f9d5fb27bc81ab4c4118a100f494f1dbb753ab6fbcb48313fd90a8169ba9aa43

Observation 63c0efb2-e7cf-4e25-958c-12399825e41c · outbound

This paper cites drink",.

TIIF-Bench: How Does Your T2I Model Follow Your Instructions? drink",

Reference 47

Resolution
verified exact
raw_fallback, observed 2026-08-07T11:33:20.978077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:33:20.868273Z digest=sha256:3e35cf40765c1aa274d7039ba913efaa1a670044b1933c06a7577d745a8b7f2f

Pith citing papers

Observation 31c843ec-e643-4c38-9c8d-3749a0c7960a · inbound

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation cites this paper.

Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T20:44:16.332457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:44:16.332457Z digest=sha256:33db794a14fdf9eb1fdf50ca76bce0046c0ac702208ed40efa95a56e3dc3b8ff

Observation ea4993d5-ebd1-4f8d-82ae-6dd76d621ebe · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T14:08:36.801359Z digest=sha256:9f2e5a826ace1fc4e3c7ae1a6b2865be7f1ee765e256cf705e3556121302bb68

Observation c225f343-de65-4a46-a361-411e08e52309 · inbound

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer cites this paper.

Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T19:47:32.879231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:47:32.879231Z digest=sha256:f36cebc1eb83dc45d58f6df60aed9e55dda8b9082b40914989f960e658160f6f

Observation cf3cbdbc-5ce8-4b02-be27-9db081e56ed6 · inbound

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models cites this paper.

PhyDetEx: Detecting and Explaining the Physical Plausibility of T2V Models TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T17:57:57.263574Z digest=sha256:240045e8fe90fe7e1ed1bec4932397352b0e5455b2e7d8f513ccc97fe40fc171

Observation 5fcee97d-a73a-42b4-a135-475efdaff784 · inbound

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition cites this paper.

MICo-150K: A Comprehensive Dataset Advancing Multi-Image Composition TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:20:58.483350Z digest=sha256:8e9178ed62b5f428726ec546361b8f7fd812beed25e52d95d40425530addbb9e

Observation 79018805-f4c4-47fb-9475-6dbc16564c8a · inbound

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models cites this paper.

DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T13:54:00.141439Z digest=sha256:9dd75d1a8a1c0966996f5a94a9056c228dbc2e5cc0a1bf124de5d9cb060fa253

Observation 5426a6ea-7c72-416f-9124-b465d34fd5ea · inbound

From Pixels to Concepts: Do Segmentation Models Understand What They Segment? cites this paper.

From Pixels to Concepts: Do Segmentation Models Understand What They Segment? TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:12:38.800119Z digest=sha256:633d19360658c084cfa3b562f97767ef9ad5fcb1e5c82764c3516eb2430352ad

Observation 90440ac1-aca8-4a53-aab1-abfff027297e · inbound

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture cites this paper.

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T05:12:37.339084Z digest=sha256:1fdb1066f8f1cdf52f38771190c225de4fbe818c408dead66f276d3049e0a0bf

Observation 10be8eba-096f-4f53-8197-1919ab7cff08 · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:11:23.775843Z digest=sha256:98668dfa11733f1cd788368eb19ba4677a12d47c7fd7732ede83d71b97d6f8f6

Observation 20830bd0-77c1-4d1b-a2ab-8587b17b6857 · inbound

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment cites this paper.

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T09:19:39.848194Z digest=sha256:f75a842afc22fa0b905ce2fef79e231fe2dbe91c0f8ac0aa95b54321536b62ba

Observation 931927c3-f230-43b4-90e6-27b7cd753e9a · inbound

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation cites this paper.

Qwen-Image-Bench: From Generation to Creation in Text-to-Image Evaluation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T14:05:25.988619Z digest=sha256:ba67a7bd81250d3a41a2e24ad26d8456254934ecd8679a418b8e7b3b37e8029f

Observation 6ce8cf1b-5b0b-4e38-9c7e-8c2f33925304 · inbound

Qwen-Image-Flash: Beyond Objective Design cites this paper.

Qwen-Image-Flash: Beyond Objective Design TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T11:00:04.775189Z digest=sha256:3d838b4a06c14e93dbf7e14d2f45ee60155f6e762bdeaa0cd8c278b84ba0e95a

Observation b904e49c-599d-4d83-bc9a-2b7878c35253 · inbound

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization cites this paper.

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T17:54:09.656061Z digest=sha256:185ea34092a4be3ca89e74c246b532441e14ac4ec4b10f6e0f825234a953aa17

Observation a05b04df-904d-4a8c-9373-77a5c65176cc · inbound

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation cites this paper.

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T00:19:49.071495Z digest=sha256:5fcd63bb314c76c952a3b420a15e9870770ddf5d6e99342eba10b058ee149630

Observation da00e805-41d0-4811-b20c-ba74f1783c8a · inbound

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist cites this paper.

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-14T01:20:44.199458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T05:15:40.555929Z digest=sha256:4d79e0c742f2d4eaa661ef4835e2210310b0660b845b7d953ee6506eedacf9e5

Observation bd85c706-22c2-4052-b818-13843b71db60 · inbound

DynEval: Holistic Evaluations of T2I Generative Models in the Wild cites this paper.

DynEval: Holistic Evaluations of T2I Generative Models in the Wild TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-14T06:20:23.251121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:20:23.251121Z digest=sha256:c1e3005fb4ed3d6bbd9b0c398980e09d401ba5387aa29430cc510551fb0f94e2

Observation 4feca48a-f4ad-44a7-a0c2-3b5d4199eb00 · inbound

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing cites this paper.

Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing TIIF-Bench: How Does Your T2I Model Follow Your Instructions?

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T13:39:02.289063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:39:02.289063Z digest=sha256:6a84b1199f5b7272245a0d45b4a2c051679c0326ee724b6e657866c6acba488f