Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:01:57.121114Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 2 inbound Pith citation observations for arXiv:2501.03059.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:01:57.121114Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:19:17.425943Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T00:05:51.755490Z
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a79eec42-9284-4ab7-b538-b06d8183c268 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Stochastic Interpolants: A Unifying Framework for Flows and Diffusions
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53b577ea-56c4-4119-b4a2-dea4fddab751 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Latent-Shift: Latent Diffusion with Temporal Shift for Efficient Text-to-Video Generation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 026bdaca-e3e6-419f-bedd-d90ba08bcb12 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Spatext: Spatio-textual representation for con- trollable image generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6545e6c9-05f6-45cc-b5ff-d5e1d6aec3b7 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Improving image generation with better captions
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e61b7d5-daca-40de-a087-2c92a7c4e001 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Understanding object dynamics for in- teractive image-to-video synthesis
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3068459e-e59b-4a28-883a-7bf4ccd5a9f3 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Stable video diffusion: Scaling latent video diffusion models to large datasets, 2023
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb540953-09a0-4a6d-9f82-963835373585 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Align your latents: High-resolution video synthesis with la- tent diffusion models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4d8372c-59c9-43c0-89a9-54498e30b8bd · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Instructpix2pix: Learning to follow image editing instructions
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d176fb41-f0cb-4dae-bb5b-44aed1de8e5c · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Videocrafter1: Open diffusion models for high-quality video generation, 2023
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d4752850-775e-48dd-b57c-61b559bcec21 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 59486bbc-3e52-4d5f-8ad1-798f0469ce3d · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 241d01f1-fc55-4112-b15e-f50fd3eb7c06 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Animateanything: Fine- grained open domain image animation with motion guid- ance, 2023
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 339b29cf-d819-4082-a1ad-ba821192fe25 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation The llama 3 herd of models, 2024
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cf8d510e-3057-4111-bc9d-7b91f7cf6eda · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4e319397-44a7-4626-88ff-29f3d1e94af5 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Preserve your own correlation: A noise prior for video diffusion models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1465803-95ba-4758-900e-5490b1021b54 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f791340-10ec-4842-ab9f-a80fa5685c64 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Animatediff: Animate your personalized text-to- image diffusion models without specific tuning, 2024
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 794d29f6-db4c-4e8d-b549-c34e33193d3d · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Latent Video Diffusion Models for High-Fidelity Long Video Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59eca75f-cf40-44da-8858-40a590a90645 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Classifier-Free Diffusion Guidance
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a9cf77-7e3b-4d9c-ae18-32865dfe6879 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Denoising dif- fusion probabilistic models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35d94a97-3fac-43a3-b083-25da50e759fe · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Imagen Video: High Definition Video Generation with Diffusion Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb44c0d-11ac-497a-b775-f095f605d541 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Video dif- fusion models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bf0d250-e71a-4f02-91e9-a6dde78bcb7c · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Auto-Encoding Variational Bayes
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e2dcf34-dc34-409c-9fd3-2beb71745ce1 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd5142c5-1752-4337-8487-7e606761e9d2 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Gligen: Open-set grounded text-to-image generation, 2023
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ba225d87-1e3a-4e6f-bf43-f36c23d0f811 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Common diffusion noise schedules and sample steps are flawed, 2024
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b4a4f9be-6999-4bef-8900-5f709a9707fd · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 465be81b-491b-4b76-a33d-52988a93c56e · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Rectified Flow: A Marginal Preserving Approach to Optimal Transport
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e1a02e-49a8-4d0f-a258-a86e57f40f0b · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c944c83-5a2c-4b2b-ba12-af961d62b9c7 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Cinemo: Consis- tent and controllable image animation with motion diffusion models, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2b2397d9-42a8-460d-a8d7-1570d3ddb895 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Latte: Latent Diffusion Transformer for Video Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e2167d8-cb15-4b8d-954a-04158ab401ab · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Snap video: Scaled spatiotemporal transformers for text-to-video synthesis
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 25e0d8ad-9146-4f52-8e0b-26e12dd0aa7d · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Compositional text-to-image gen- eration with dense blob representations, 2024
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bbdd92b7-fdd3-4aad-b416-20b7ce96c201 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Video generation models as world simula- tors
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 89497550-e4ca-4777-808e-45c596c84c8b · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Video generation from sin- gle semantic label map
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 87207992-9770-43ec-9487-d23ab702bce9 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Scalable diffusion models with transformers
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a42d2a5-590c-46f4-beeb-241a526e16e5 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd62e78-f0a7-4dc4-9b9e-005dbf124fe3 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Sampson, Shikai Li, Simone Parmeggiani, Steve Fine, Tara Fowler, Vladan Petro- vic, and Yuming Du
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f453eeed-d4f3-4658-be80-4da6ce43f96f · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Learning transferable visual models from natural language supervision, 2021
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20d15cec-fdb4-4863-997d-6ba41dcf599b · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Sam 2: Segment anything in images and videos,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a0e9035-f195-4696-ba21-d4e43d5c5ee0 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Consisti2v: Enhancing visual consistency for image-to-video generation, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e05b1915-93e3-4c75-9963-b40e97d5e580 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation High-resolution image synthesis with latent diffusion models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 660f2410-fdf5-4973-b83a-6a28e00e162a · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling, 2024
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7c554da8-0fcd-4a0f-b428-c38a9cd680d0 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Make-a-video: Text-to-video generation without text-video data, 2022
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 556750ab-c2a3-4f8c-91aa-0055e0b78c60 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Denoising Diffusion Implicit Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0bf9f6c-004e-4650-9f91-47a4fd3ef298 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Score-Based Generative Modeling through Stochastic Differential Equations
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb07f7b0-7508-4810-994a-3049dd0e15da · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Raft: Recurrent all-pairs field transforms for optical flow, 2020
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 979d90b6-12b3-4aa0-b422-001609b75440 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation To- wards accurate generative models of video: A new metric & challenges, 2019
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 152c2249-a9ce-4d79-8163-89c3773a278d · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation ModelScope Text-to-Video Technical Report
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0060c327-9d8f-41d8-adbe-7add5337125d · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Swap Attention in Spatiotemporal Diffusions for Text-to-Video Generation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e504a121-af29-4c76-9b00-cb055b0f80da · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c74e4b38-fd33-4197-b3c8-0f0724d36a24 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Internvid: A large-scale video-text dataset for multimodal understanding and generation, 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a6138c38-078e-4e60-9570-b216a2aeeaea · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Cvpr 2023 text guided video editing competition, 2023
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e80e4cc-1dec-4c74-a97f-0de74b0da988 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Dynamicrafter: Animating open-domain im- ages with video diffusion priors, 2023
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ee53ac5c-e98f-45e2-81dc-a4a741fb77d6 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation I2vgen-xl: High-quality image-to-video synthe- sis via cascaded diffusion models, 2023
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation f2cdff7d-03cc-484b-98ba-aca83a4d4f89 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation MagicVideo: Efficient Video Generation With Latent Diffusion Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45863543-3017-4676-844a-782aeaa0ff15 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation Qualitative Comparison of Masked Attention Mechanism Fig
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2c53be84-e420-4376-9e28-fbfd94990343 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation 3.1, our pre-processing pipeline ex- tracts a motion-specific prompt, cmotion, from the input text c, using a pre-trained LLM
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a1db1218-8197-43f9-ba37-0f12a174fafb · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation 3.1, the pre-processing process be- gins with extracting motion-capable object prompts from the global prompt c
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4ca3f9d3-c478-45b9-890a-0d95840fbab9 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation First, the initial segmenta- tion s(0) is extracted from x(0) using SAM2 [40]
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9ed908a2-eb82-4c35-8cdd-cba69f8a898a · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation The first is the U-Net architecture
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation b01e8fac-5167-4c7d-9b03-3d4a8a2b7c53 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation The filtering of 128 videos, out of the full SA-V dataset, involved several steps
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 16d604f0-777d-47f1-8ccf-fb2d0fb1ceb4 · outbound
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation description of overall motion
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 988a81d2-fcf8-43f8-b760-b76e2c8df26f · inbound
Seeing Voices: Generating A-Roll Video from Audio with Mirage Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e46e91-26c5-42ff-a16e-a8bcfdf52b99 · inbound
Evolution of Video Generative Foundations Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
Reference 227
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.