Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T23:11:58.580004Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2605.31590.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T23:11:58.580004Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 76e9be83-1093-4ba6-8b1b-1331471cb602 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6203fcb5-ab24-4cf3-a47c-1cebb3ebf072 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Ditctrl: Exploring attention control in multi-modal diffusion transformer for tuning-free multi-prompt longer video gen- eration
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d906fdfe-4d9d-45bc-ab08-b0d23ac89454 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Ditctrl: Exploring attention control in multi-modal diffusion transformer for tuning-free multi-prompt longer video gener- ation, 2025
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d66cb7dc-8fd3-409e-8000-5e30508060aa · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Videocrafter2: Overcoming data limitations for high-quality video diffusion models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b51b9e38-4e89-40b8-a1af-5a3718d22875 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Training-free layout control with cross-attention guidance
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d793ee3-ce5e-4d34-880d-8f05ac37538a · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gentron: Diffusion transformers for image and video generation, 2024
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af1d1ea9-6c2d-4183-b263-ee3a56eb75d0 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation arXiv preprint arXiv:2403.05131 (2024)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a7f9ca58-707a-473f-b164-874037155d13 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gemini 2.5: Pushing the frontier with advanced reason- ing, multimodality, long context, and next generation agentic capabilities, 2025
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edb8bad0-c096-4c02-8693-3e28f47d70ea · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Worldscore: A unified evaluation benchmark for world generation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3a84c555-3e4e-45cc-8807-04266b9a91e6 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a691235-a635-4b97-bfc6-be22af53c68d · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 11862399-67f1-4211-9819-3e3d43bd61b1 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Vbench: Compre- hensive benchmark suite for video generative models, 2023
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b5af13c-43e8-453f-ae57-5aa4f38c006c · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Vbench: Comprehensive benchmark suite for video generative models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 765ac7fd-7f41-4451-9ee2-d308f73c225a · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Vbench++: Comprehensive and versatile benchmark suite for video generative models, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d0dd1e-57da-4b78-9f0b-8259decb59f9 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Ultralytics yolo11, 2024
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca0989ee-7986-43b7-afd5-4559bbd3073d · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation How Far is Video Generation from World Model: A Physical Law Perspective
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 54b066d2-8886-4daa-8188-e8604c52732a · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Shotadapter: Text-to-multi- shot video generation with diffusion models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88f5ef1e-8b26-4a05-96cc-bd7cf70ee4ee · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Fifo-diffusion: Generating infinite videos from text without training.Advances in Neural Information Processing Sys- tems, 37:89834–89868, 2024
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c780c3ee-e856-491d-a579-0b1aa5b91ff0 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation WorldModelBench: Judging Video Generation Models As World Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 05f890ee-2a95-4b28-86ee-f2fee442a67a · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation WF-VAE: Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dfe4843f-5617-403d-9aad-ce30152082b3 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f0e805ef-54a3-4379-bebe-ef4974e53807 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Videoinsta: Zero-shot long video understanding via informative spatial- temporal reasoning with llms
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bfb109e-91e9-485e-9ee6-7c1c3ce1ce81 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gentkg: Generative forecasting on temporal knowl- edge graph with large language models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa1e13c-f5a5-47af-8f0f-e4f841d6fa32 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation When and where do events switch in multi-event video generation?, 2025
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 103510e3-8a9d-4276-a83a-94fcbaeb3d46 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Open-Sora Plan: Open-Source Large Video Generation Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1ed6f011-4c01-4cd5-9f57-32d103dfb18b · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation WorldWeaver: Generating Long-Horizon Video Worlds via Rich Perception
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2946d51b-53f8-42b6-b0d1-87d41cca2d5b · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation The shape variational autoencoder: A deep generative model of part- segmented 3d objects
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 710f9f24-21a5-480d-8620-2ab216540981 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Mevg: Multi-event video generation with text-to-video models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e3ae73-d0be-467e-9399-74ddc6fc4b33 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Dinov2: Learning robust visual features without supervision, 2024
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52bacbbf-0c80-4848-b06d-17dd0bf0e361 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Scalable diffusion models with transformers, 2023
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9afdd7-eae1-4c8f-8833-a4f50f5de467 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Movie Gen: A Cast of Media Foundation Models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 36c43547-2fd7-4901-9b64-9422f20836ac · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Maskˆ 2dit: Dual mask-based diffusion transformer for multi-scene long video generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a8f765a-cdf3-4b4a-81dc-82c23a4612f7 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Mask2dit: Dual mask-based diffusion transformer for multi-scene long video generation, 2025
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08d7fbad-a0db-4e59-8deb-80dfb991969a · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Freenoise: Tuning-free longer video diffusion via noise rescheduling, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aceb53e7-9a54-4806-9764-686d6efe35e8 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Learning transferable visual models from natural language supervision, 2021
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b179efed-ebcd-47df-9319-4be0b7ace8a5 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Sam 2: Segment anything in images and videos, 2024
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f77088c5-2ede-446d-8ef1-533c7849f217 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Sentence-bert: Sentence embeddings using siamese bert-networks
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4aae7ec2-705b-461d-82a0-9946e25a089c · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Hunyuan-Large: An Open-Source MoE Model with 52 Billion Activated Parameters by Tencent
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ea07cecb-0ba7-432c-b451-300646f8e850 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Mochi 1.https://github.com/ genmoai/models, 2024
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0817d67-ba4d-4308-bdd9-13311526e164 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation The tensor brain: A unified theory of perception, memory, and semantic decoding.Neu- ral Computation, 35(2):156–227, 2023
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 058e16f0-6d4c-47a7-bb77-7dc14683e4c3 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Wan: Open and advanced large-scale video gener- ative models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b2d46f0-9827-4cb1-9f7a-9c650d46af2c · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d81d4a1-cccb-442e-a445-82911ff79c1f · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Gen-l-video: Multi-text to long video generation via temporal co-denoising, 2023
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9812266b-e47a-46e7-b7ba-f5981c6d82f3 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Internvid: A large-scale video-text dataset for multimodal understanding and generation, 2024
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49b15346-04e9-4de5-ba0f-0811e85e1e5c · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Mind the time: Temporally-controlled multi-event video generation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ce5e9c2-80d1-4d58-84f8-94c3b40d366f · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation A survey on video diffusion models, 2024
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398bc0c1-7a70-4f80-a3e8-ca4200f8a894 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Dynamic prompt learning: Addressing cross- attention leakage for text-based image editing.Advances in Neural Information Processing Systems, 36:26291–26303,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deef2a6c-e11d-468c-968e-25a8d71a2339 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 60284256-0f1b-4488-ad4c-fa16fe09a7ca · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Instadrive: Instance-aware driving world models for realistic and consistent video generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9625d3ce-b586-4fa2-a39b-9289a657270d · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Evaluation Agent: Efficient and Promptable Evaluation Framework for Visual Generative Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2a0daf92-4b15-4ffd-9ac8-0a55efa4008b · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Magiccomp: Training-free dual-phase refinement for compositional video generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eb376d5a-e557-4652-bfa1-cdf7be0c9a7b · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Vbench-2.0: Advancing video generation benchmark suite for intrinsic faithfulness,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4d837c3-d93d-4010-bbf0-90d83ba5839e · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6235ece4-5e82-4d51-b6f2-d3b9ca5d911e · outbound
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e687d6ff-4959-4311-a7cc-fff26b7fe478 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 449ade3c-c51f-44b1-8b62-8659f9268fd1 · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9e9fd10-a838-4c3e-bcca-32d14d6c02aa · outbound
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Video Generation TunerDiT achieves superior Text-Video Alignment scores compared to the base- line models, including open-source base models and zero- shot methods MEVG, DiTCtrl and FreeNoise
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.