Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:17:38.923422Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 3 inbound Pith citation observations for arXiv:2505.07652.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T22:17:38.923422Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:25.127558Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:19:30.438596Z
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 13678533-e14c-4dc6-935e-74d8b7ff31ef · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Kling ai
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bff35877-1525-4efa-bada-1c9879acd1e3 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Universal guidance for diffusion models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0927c5f7-50fd-4bd9-ba9e-cb74ca6b6720 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Pyscenedetect: A cross-platform tool for video scene detection
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 421f36e9-7aeb-42ed-93a8-69c5bd344957 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Video generation models as world simulators, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation daf16dce-778a-4038-b148-88eff605474e · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Emerg- ing properties in self-supervised vision transformers
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bac505e7-2db1-4e76-8819-3cbd58540c02 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Scene detection in videos using shot clustering and sequence alignment
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 150b4bf3-da9d-4f53-8171-b9184be0d4f2 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Gentron: Diffusion trans- formers for image and video generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3147f4f3-22b1-40a4-af91-bed7f87fb9d1 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Seine: Short-to-long video diffu- sion model for generative transition and prediction
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bb04a63b-a8c9-476a-a408-095a31f43532 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Efficient video prediction via sparsely conditioned flow matching
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc3e2f14-21dc-40a1-b278-4312ac90f489 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Arcface: Additive angular margin loss for deep face recognition
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebaf899a-4e70-4c44-bc48-84c366975214 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dfc15bb-947d-417e-852d-371e2d5cbf0e · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models An image is worth one word: Personalizing text-to-image gener- ation using textual inversion
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 67c0e663-dbd0-4206-b3a5-9dfaa0179164 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Preserve your own correlation: A noise prior for video diffusion models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3683d5e7-e7f6-419d-96d6-26702eda47e0 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Tokenflow: Consistent diffusion features for consistent video editing
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5f047eb8-026d-4dd3-b7d0-09eb4260100f · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Photorealistic Video Generation with Diffusion Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8042d70-fc6c-4ce7-99ae-5f167a73af98 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 484d5403-2b65-4494-a126-619a26ae8f65 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Style aligned image generation via shared atten- tion
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e3abdbb-6f5d-47e6-90a2-46a8fc963656 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Denoising dif- fusion probabilistic models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e4923c-24c5-4724-9a51-eb4f449bc3f3 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Vbench: Comprehensive bench- mark suite for video generative models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3db3973c-c3ff-417f-bf7d-a4645ef8a3dd · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Pyramidal flow matching for efficient video generative modeling
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c886c939-06cd-4e18-9767-10277bb7470e · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Rave: Randomized noise shuf- fling for fast and consistent video editing with diffusion mod- els
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffaa91db-4b72-4f0b-82a6-8068d7d324ca · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Text2video-zero: Text- to-image diffusion models are zero-shot video generators
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 01e55a88-6fb6-4935-bb92-07feccdbf8a5 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Open-sora-plan, 2024
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd21cfb2-33d6-41fc-bd41-9e0708e3d24a · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Dream machine
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b0cfaac7-6e24-4102-a408-e8f56f8f01bc · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Photomaker: Customizing re- 9 alistic human photos via stacked id embedding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d250b7e7-d8bb-4f4f-87ab-97dc96d1378a · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Decoupled weight de- cay regularization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2d6245f-540e-4357-8b26-9f9587b02e38 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Subject- diffusion: Open domain personalized text-to-image genera- tion without test-time fine-tuning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feac9aea-0d51-4339-b06f-727199b10015 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Latte: Latent Diffusion Transformer for Video Generation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72bf8f92-7891-423a-b888-1e303b04b61f · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d1fcb8bf-6d48-41dd-94df-0fe893535e37 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Mevg: Multi-event video generation with text-to-video models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation edadf137-e43e-4fb7-b4cb-1e06cbab4294 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Chatgpt: A large language model
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0fdf758a-16f1-48ad-aced-b2372144e029 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Scalable diffusion models with transformers
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b488866c-0024-415c-bfe0-813538cae67b · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Sdxl: Improving latent diffusion models for high-resolution image synthesis
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b0bb7b0b-f70f-415b-9b0e-8c7fe2035584 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Movie Gen: A Cast of Media Foundation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d019136f-e93c-43df-8e57-cb4d109af167 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Prolific: Online participant recruitment for surveys and research
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 44b3f07b-f08d-4217-ac57-ed0ca612e5e5 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Freenoise: Tuning- free longer video diffusion via noise rescheduling
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation cee252ca-5c9b-46ca-a96e-d125a2a252bd · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Hierarchical text-conditional image gen- eration with clip latents
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6f801bd2-c66e-4cde-8115-4557a034eb76 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models High-resolution image synthesis with latent diffusion models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf21d14-3ea7-4613-a044-388d7225c89d · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2dd2e66-6f3c-42b4-9e36-36a9debae172 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5b775e97-b644-47b1-b73d-2a409e39f3ff · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Make-a-video: Text-to-video generation without text-video data
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3ed16200-56e0-48bb-8ddd-40a84f0125f4 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Deep unsupervised learning using nonequilibrium thermodynamics
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13cf3a9a-5757-45ec-acd0-c7ac268f91c3 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Raft: Recurrent all-pairs field transforms for optical flow
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61ea114c-61e2-4299-a2a1-ec159051aff1 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Training-free consis- tent text-to-image generation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 66e89b28-7241-4d84-a48f-f52dc5e841ec · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Film history: An introduction
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9e45f1bc-2c20-4267-be71-32d6c7bf78b2 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models YOLOv5: A state-of-the-art real-time object de- tection system
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c52dfaaf-5e98-4ae7-86e6-73bb32b5ab12 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models FVD: A new metric for video generation, 2019
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 46ad2349-9303-40b9-9ecc-a72f9ab648bf · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Gen-l-video: Multi-text to long video generation via temporal co-denoising, 2023
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1264697e-8fd4-46e2-8a1c-30a905398aa9 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models InstantID: Zero-shot Identity-Preserving Generation in Seconds
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d4b6c6-58bb-4996-9253-29a8e56b508d · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Internvid: A large-scale video-text dataset for mul- timodal understanding and generation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 073c7195-f5bd-4f29-aaf0-3f39537a7587 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Elite: Encoding visual con- cepts into textual embeddings for customized text-to-image generation
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0740253b-cf97-4a28-9174-c46137e68821 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ad1eeb89-1b14-46ad-9986-218bc092ba9d · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Fastcomposer: Tuning-free multi- subject image generation with localized attention
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79302690-86fc-43e8-acf0-d8f658a8e940 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models FaceStudio: Put Your Face Everywhere in Seconds
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14bf39c8-b05f-4c27-a6de-e0c5299436e6 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b2e5d4a-ef95-4327-abeb-dc682749d03f · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Automatic partitioning of full-motion video.Multimedia systems, 1:10–28, 1993
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d40f09fb-7899-40a4-a32d-11fa6831784e · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Llava-next: A strong zero-shot video understanding model
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2f4eead7-6685-4040-84ba-bccb3d848609 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Layoutdiffusion: Controllable diffu- sion model for layout-to-image generation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e2160dad-e90f-42e3-bf05-8d3954f2ed2f · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Open-sora: Democratizing efficient video production for all
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 17f2e3ad-8568-4fbf-b70d-01329b46c249 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8132c316-0c14-4c3f-936f-4124ee2c06df · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models StoryMaker: Towards Holistic Consistent Characters in Text-to-image Generation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4ed17f5-8423-4708-9276-6337c466ddb3 · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models Unresolved cited work
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 29608644-589f-48f6-afcf-c0b995dace0d · outbound
ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models a man reads a book under tree
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 25fc22a4-82b8-4e38-a725-ab4140fa5bcc · inbound
LoViC: Efficient Long Video Generation with Context Compression ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91725449-9114-4041-b4f5-ecb4f5dd7b41 · inbound
GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 392e4b67-0a10-4b7c-8b45-0e3d49bc1103 · inbound
GroundShot: Visually Consistent Multi-Shot Long Video Generation via Entity-Grounded Shot Scheduling ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.