Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 49 inbound Pith citation observations for arXiv:2409.10695.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T01:00:03.278624Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T05:54:33.573861Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 980e8f8d-e3ac-469e-800e-3510a6600510 · inbound
Emu3: Next-Token Prediction is All You Need Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 07ee6c8e-8d94-4799-a876-ac6f425c9e61 · inbound
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformers Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2977e68d-84b7-41bd-ada7-f78dfcb54bca · inbound
$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 01114a68-abbe-4329-91d7-036312b2936f · inbound
Open-Sora Plan: Open-Source Large Video Generation Model Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 355e14f9-495f-4100-942e-305553f47ac2 · inbound
IQA-Adapter: Exploring Knowledge Transfer from Image Quality Assessment to Diffusion-based Generative Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 142d523d-929d-47ef-826e-9084208b1384 · inbound
X-Prompt: Towards Universal In-Context Image Generation in Auto-Regressive Vision Language Foundation Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a114ed54-6af2-4595-9c78-ef507dfa1625 · inbound
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68d6a65-9d14-42f4-86ce-3d4ff418d0bb · inbound
Chimera: Improving Generalist Model with Domain-Specific Experts Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdcbe4cc-da18-4798-bece-32dd43acc6ff · inbound
EasyRef: Omni-Generalized Group Image Reference for Diffusion Models via Multimodal LLM Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8ebef6-b50d-4aee-ab40-3d7f7d1f899e · inbound
SnapGen: Taming High-Resolution Text-to-Image Models for Mobile Devices with Efficient Architectures and Training Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a64fa0b6-6fba-4423-8591-c0e820380f6f · inbound
Autoregressive Video Generation without Vector Quantization Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2b7d4d9a-5f22-48e9-8f3c-87bf35b52059 · inbound
1.58-bit FLUX Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a4f5851-5d71-4b49-9184-deeaefb6e187 · inbound
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a02f3cc-7f46-4de5-8ec3-04ece55250d7 · inbound
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dac03e54-6e18-4ed8-ae35-d7cff5977a2c · inbound
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39954511-0346-4579-987d-00fb6dd8ff98 · inbound
MSF: Efficient Diffusion Model Via Multi-Scale Latent Factorize Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6399c26-8129-41ab-b7d8-37bb20ad8fcb · inbound
Diffusion Generative Modeling for Spatially Resolved Gene Expression Inference from Histology Images Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e88ed9b2-121a-47e5-9f2c-726b37ef1b9e · inbound
SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ed10bf5-00e1-4e10-b85f-b2740dd74a1c · inbound
I Think, Therefore I Diffuse: Enabling Multimodal In-Context Reasoning in Diffusion Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f33ea93-ae88-49b2-a526-072b95ab80fc · inbound
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b7f518a-5f1d-4917-824d-cf9e5d79b6bf · inbound
RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c9985a9-f60c-436a-9496-d462c89a3dda · inbound
Normalized Attention Guidance: Universal Negative Guidance for Diffusion Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25002f22-4f2d-4fbf-9222-f49370bb2c88 · inbound
RICO: Improving Accuracy and Completeness in Image Recaptioning via Visual Reconstruction Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 715e4c1a-801f-4b0a-ba64-1693121f496e · inbound
VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09a834a-4aeb-4896-b560-eb95b88b368c · inbound
Ultra-High-Resolution Image Synthesis: Data, Method and Evaluation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c750e13-1c5c-4021-aa34-6f2c99e27283 · inbound
UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 78c10748-bbdd-4373-b842-1ba1ad57c327 · inbound
A Comprehensive Study of Decoder-Only LLMs for Text-to-Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d796aa71-2a65-4bfb-8a60-b5db591ac07a · inbound
PosterCraft: Rethinking High-Quality Aesthetic Poster Generation in a Unified Framework Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e9cc49-9856-4f30-86c5-0eab0a8f9263 · inbound
CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a3d0dcb-6129-494c-8159-78cd2e95eefa · inbound
DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8ca69b2-8667-48f6-b93e-685d041ce13e · inbound
Towards Evaluating Robustness of Prompt Adherence in Text to Image Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43827c62-29e4-4125-abcf-d0cb4571445f · inbound
PoemTale Diffusion: Minimising Information Loss in Poem to Image Generation with Multi-Stage Prompt Refinement Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dfb72e3-ca11-4f38-bf33-48afc8e80cc3 · inbound
Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2799051a-0be3-4f43-b590-fc8bc919b5c3 · inbound
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a3cf6d0-8645-4517-8060-6e46deb526bf · inbound
Reconstruction Alignment Improves Unified Multimodal Models Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a9c309-62ac-4231-b2ca-f588b283ca0e · inbound
PixelDiT: Pixel Diffusion Transformers for Image Generation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 66178e1a-860b-4bc5-8d26-ebf20949ed5e · inbound
LTX-2: Efficient Joint Audio-Visual Foundation Model Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 32dab576-631f-471d-9d62-61a173be5449 · inbound
SnapGen++: Unleashing Diffusion Transformers for Efficient High-Fidelity Image Generation on Edge Devices Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa2482c6-3010-442f-898e-3e3ff850af99 · inbound
WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81dce721-b284-41cc-8f68-c2ec56106a31 · inbound
Self-Adversarial One Step Generation via Condition Shifting Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6cdca740-2263-4059-b665-c36a3b372e93 · inbound
Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representation Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 32535488-c446-47fe-a0f8-0ef98a2b0971 · inbound
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c1a8333e-c3ff-4c31-9b29-2753e52145dc · inbound
PixVerve: Advancing Native UHR Image Generation to 100MP with a Large-Scale High-Quality Dataset Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 08d28f16-20af-41e7-9b0c-9632f7f59757 · inbound
Rethinking Cross-Layer Information Routing in Diffusion Transformers Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dedc7eea-de42-41a8-a0c0-99e448820e77 · inbound
Rethinking Cross-Layer Information Routing in Diffusion Transformers Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1fa295a3-67e7-41c1-b558-572677447c3a · inbound
MRT: Masked Region Transformer for Layered Image Generation and Editing at Scale Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 768b2830-16b2-463c-95a0-c5816281b7a7 · inbound
OctoT2I: A Self-Evolving Agentic Text-to-Image Router Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 45069593-a195-4503-8515-5c61b7404054 · inbound
Token-to-Token Alignment of Text Embeddings for Semantic Blending Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e9ae30cd-31e1-49b4-b55f-ea999ddbf0a9 · inbound
Analysis-by-Proxy: Localization Signals in VLMs Operating as Condition Encoders Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.