Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:11:00.639240Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 2 inbound Pith citation observations for arXiv:2412.12391.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:11:00.639240Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T20:30:31.258053Z
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6eb7e9bf-7027-4ae4-9872-dd1ec108ed8a · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation [Online; accessed 4-March-2024]
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 65abf0c2-4b3b-4339-b2be-ae7f0ca877a8 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb0157df-e0aa-498e-9542-d9fb1d151351 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Mistral 7B
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df311dd2-2ef2-449c-9626-236fec9ce460 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Scaling Laws for Neural Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb023b8d-cfb7-4994-8ec8-83b8fd4eff4d · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3028d1b1-2011-410d-8a7c-5f5fa92057f3 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09efea52-30c3-46c0-9fc2-2064026aefee · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14bd2e94-787a-4f10-8185-6e0c67911dc0 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation U-net: Convolutional networks for biomedical image segmentation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7df0b84d-929d-409a-8743-63c36ea9c0a2 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 710f86f6-d71a-4824-8142-9e6a882a290c · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df250891-c198-42ef-8cf7-c91b44fff2f3 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a17937c-00ab-4842-8bb1-8e8b85577898 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699907c1-0d15-4c46-a919-6f33a9bda6c4 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da66e0fe-1c08-4771-96c4-71e9a005ca97 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation LensArt consists of 250 million image-text pairs, carefully selected from an initial pool of 1 billion noisy web image-text pairs
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5ba47946-b5f7-4d59-a311-a52ff46defe1 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation B.1 B ENCHMARKS We outline the evaluation benchmarks used to assess the performance of our image inpainting and canny edge conditioning models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b2474e7-bbe9-4e35-b4b1-4e433df4a3e6 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d2d2f455-3e5b-43d7-beaa-68fe9dad98db · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation In addition to TIFA and ImageReward, we also provide the FID score, which measures the fidelity or similarity of the generated images to the groundtruth images
Reference 2014
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 001bea12-8fd3-4ba3-bce3-ce49bc74b0c7 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation LLaMA: Open and Efficient Foundation Language Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30cca4b0-6871-4574-a296-a74f49fb3437 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation TIFA: Accurate and Interpretable Text-to-Image Faithfulness Evaluation with Question Answering
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99a30b2d-0b47-440f-911c-a8e9c81fdd8d · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation Denoising Diffusion Implicit Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e79e9040-d0fe-4f84-88b8-78fb9ef02d14 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation LoRA: Low-Rank Adaptation of Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d23cdeee-dccc-42e5-91c7-345a749c17a9 · outbound
Efficient Scaling of Diffusion Transformers for Text-to-Image Generation On the Importance of Noise Scheduling for Diffusion Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e89e2380-e22a-43ab-8d62-db81ad126e7e · inbound
Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52761174-4d24-43fc-a1f9-7293d6313353 · inbound
Importance-Aware OBS Pruning for Diffusion Models Efficient Scaling of Diffusion Transformers for Text-to-Image Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.