Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T22:55:44.984216Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 21 inbound Pith citation observations for arXiv:2509.06945.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T22:55:44.984216Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:44:28.524092Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:39:58.253188Z
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5124f506-35d6-405b-8e95-25bfe6394496 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10576f2d-416c-403f-a149-1fc83f13bcee · outbound
Interleaving Reasoning for Better Text-to-Image Generation Emerging Properties in Unified Multimodal Pretraining
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f56752a-f678-4670-b871-473500b61ff1 · outbound
Interleaving Reasoning for Better Text-to-Image Generation GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f0d6a40-9d29-4a19-8936-10f296ef8d0b · outbound
Interleaving Reasoning for Better Text-to-Image Generation SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44623535-7df1-419b-9f58-44f48e8dbe26 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Classifier-Free Diffusion Guidance
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9bcf84-7d04-4b07-b310-f5bb9393f428 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400b2981-4301-4a12-b296-a09abfdf8ee5 · outbound
Interleaving Reasoning for Better Text-to-Image Generation OpenAI o1 System Card
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90676a64-8725-46a9-ad1c-173bd31af4e9 · outbound
Interleaving Reasoning for Better Text-to-Image Generation T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e913c196-2937-4c88-9adc-fc01578ed526 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c94664b9-3afc-4601-b1c1-1ee5a6dfc69f · outbound
Interleaving Reasoning for Better Text-to-Image Generation Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ecae873-fb11-4561-9f18-6b0b8472fce7 · outbound
Interleaving Reasoning for Better Text-to-Image Generation UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e958345e-03c1-4134-a54f-10d17e56ae3b · outbound
Interleaving Reasoning for Better Text-to-Image Generation World Model on Million-Length Video And Language With Blockwise RingAttention
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 143acc15-20e2-4963-bea2-b262de52529a · outbound
Interleaving Reasoning for Better Text-to-Image Generation Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1de9c5af-6f41-4e77-854c-14ae561f2ab8 · outbound
Interleaving Reasoning for Better Text-to-Image Generation JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8fef04f-988b-4924-b6c2-dcdc3e59eb9e · outbound
Interleaving Reasoning for Better Text-to-Image Generation WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8417e21d-2b57-426d-bb17-b6ae00a836fd · outbound
Interleaving Reasoning for Better Text-to-Image Generation Transfer between Modalities with MetaQueries
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aea23b05-9223-43c1-b718-afbfdddcb612 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Lumina-Image 2.0: A Unified and Efficient Image Generative Framework
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6a80fd0-984a-4d3a-b633-fccfc4c216a0 · outbound
Interleaving Reasoning for Better Text-to-Image Generation TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c5b707e-5997-40d3-8563-10b5c8b89bce · outbound
Interleaving Reasoning for Better Text-to-Image Generation Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6149699a-4a80-4225-a911-7bf60ff7540a · outbound
Interleaving Reasoning for Better Text-to-Image Generation Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f3efb74-57f0-4b11-ae1b-346a7a6f4f81 · outbound
Interleaving Reasoning for Better Text-to-Image Generation MetaMorph: Multimodal Understanding and Generation via Instruction Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7a83eec-5066-409a-87c3-31ec6c7fb86f · outbound
Interleaving Reasoning for Better Text-to-Image Generation ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84579299-7ced-436f-8ae8-b816cbe7fedd · outbound
Interleaving Reasoning for Better Text-to-Image Generation OmniGen2: Towards Instruction-Aligned Multimodal Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 558594d0-d1f6-48f9-a683-2152cbed355b · outbound
Interleaving Reasoning for Better Text-to-Image Generation Show-o2: Improved Native Unified Multimodal Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573de087-0f8f-42a2-9344-7eb92521f4f1 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Qwen3 Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b87d6789-8fed-404b-bf6e-8d880ea1d689 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe2b674-3fa5-42d5-8723-e92f31c46a4b · outbound
Interleaving Reasoning for Better Text-to-Image Generation From Reflection to Perfection: Scaling Inference-Time Optimization for Text-to-Image Diffusion Models via Reflection Tuning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93875204-2588-4310-b563-7bab02756f4c · outbound
Interleaving Reasoning for Better Text-to-Image Generation LMFusion: Adapting Pretrained Language Models for Multimodal Generation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31b420a8-9e27-48f7-be8e-059af6970582 · outbound
Interleaving Reasoning for Better Text-to-Image Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8235d2a8-716d-4e57-a318-1de3f1d2d672 · outbound
Interleaving Reasoning for Better Text-to-Image Generation Advancing multimodal reasoning: From optimized cold start to staged reinforcement learning.arXiv preprint arXiv:2506.04207, 2025b
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1c9beee-3b0d-41a7-af2f-2b5562734f09 · outbound
Interleaving Reasoning for Better Text-to-Image Generation OneIG-Bench: Omni-dimensional Nuanced Evaluation for Image Generation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a4ba74c-4fae-4ac1-a9d8-8357d14b0a7a · inbound
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation Interleaving Reasoning for Better Text-to-Image Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 685c24f6-8ff2-4100-87cb-c52c064ff8e9 · inbound
How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning Interleaving Reasoning for Better Text-to-Image Generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b316f3e0-46ec-46b5-8fb3-fd08f3daa2ec · inbound
Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment Interleaving Reasoning for Better Text-to-Image Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5f51071-fd0f-44b1-a625-9e7400416a2c · inbound
TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training Interleaving Reasoning for Better Text-to-Image Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3f383230-214a-4a49-86a1-679e9d4442f1 · inbound
TorchUMM: A Unified Multimodal Model Codebase for Evaluation, Analysis, and Post-training Interleaving Reasoning for Better Text-to-Image Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 78c1b218-e674-4b05-b64c-8811eeaa877c · inbound
Meta-CoT: Enhancing Granularity and Generalization in Image Editing Interleaving Reasoning for Better Text-to-Image Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b890564b-ba3d-47d4-89d5-74e044db0147 · inbound
Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d1ab0b94-b4e9-4270-915d-cad94ddabee2 · inbound
SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation Interleaving Reasoning for Better Text-to-Image Generation
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cd069034-a83f-49d8-b350-c40ee2a3e0bb · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e91cf0d4-220b-4294-bbb2-e2c4dc308fb7 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b968cffc-d6c6-4451-a5bb-9ad3dae66c79 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d3b2a36d-e9d5-4c90-bf62-5f9a86decb77 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b10ec737-893f-4828-8850-e260f3b6d7e4 · inbound
Flow-OPD: On-Policy Distillation for Flow Matching Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2480ca98-713e-4083-877f-7bcb775e4451 · inbound
UniPath: Adaptive Coordination of Understanding and Generation for Unified Multimodal Reasoning Interleaving Reasoning for Better Text-to-Image Generation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4d94f5ca-30ea-45a2-93af-68710ce439fb · inbound
Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning Interleaving Reasoning for Better Text-to-Image Generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 28de1f0c-778d-4778-bf02-b61f1fa7d7d1 · inbound
LatentUMM: Dual Latent Alignment for Unified Multimodal Models Interleaving Reasoning for Better Text-to-Image Generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 86c8d491-79c1-4d6b-89f9-afd5f2f656cf · inbound
OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration Interleaving Reasoning for Better Text-to-Image Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2e4f6a9a-d3a6-417a-897c-15e969a7484e · inbound
Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning Interleaving Reasoning for Better Text-to-Image Generation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 66a7c912-fa2b-4992-b99e-417e7cce5358 · inbound
IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation Interleaving Reasoning for Better Text-to-Image Generation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eee4709d-d147-47d7-81fa-b1206a14b403 · inbound
SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search Interleaving Reasoning for Better Text-to-Image Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 097f2bcd-f015-46ea-b082-5db26a881451 · inbound
Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent Interleaving Reasoning for Better Text-to-Image Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.