Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:24:59.444766Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 32 inbound Pith citation observations for arXiv:2508.18621.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:24:59.444766Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T00:05:47.808077Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T14:48:32.703864Z
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4968f6d7-9c97-4a4f-b775-bba1ef82af99 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Qwen2.5-VL Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fb76f94-98a9-4e52-b57a-a60ac840fe62 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04bda69f-6eb3-4468-bb05-9d351ea780ed · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6897fdfb-9ade-43ff-a776-5b1078b0f010 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation EMO2: End-Effector Guided Audio-Driven Avatar Video Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ae816ba-4882-42a5-9c5f-4b5e41715aa9 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Wan: Open and Advanced Large-Scale Video Generative Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54d8d8a5-a960-4da1-b302-d210086694cd · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163e6c2f-8abb-466f-b17d-b21ee7be9c84 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Koala-36M: A Large-scale Video Dataset Improving Consistency between Fine-grained Conditions and Video Content
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ac0e266-00bd-4c76-9f87-ba7e83315f05 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13af52ce-5fa9-44ae-8acb-0f90b09b0c9c · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Haoning Wu, Erli Zhang, Liang Liao, Chaofeng Chen, Jingwen Hou Hou, Annan Wang, Wenxiu Sun Sun, Qiong Yan, and Weisi Lin
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca1bd54f-4053-41a3-b5f2-245a5bca4ded · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Unresolved cited work
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7516b7e-1033-429c-b33c-0060dd4bfa65 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Christoph Schuhmann
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b29b68cb-f0fc-44d7-9b16-1208154a8507 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Image quality metrics: Psnr vs
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d77ea705-6962-4e0f-b2a1-7babb4f8852b · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation ViTPose: Simple Vision Transformer Baselines for Human Pose Estimation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90c58207-0cbf-4150-bb93-e2f38d6b30c1 · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation Flow Matching for Generative Modeling
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d377db72-a42b-49dc-a21b-bcd2603004dc · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation USP: A Unified Sequence Parallelism Approach for Long Context Generative AI
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e091873-51f3-4079-b666-ccbce3d342bf · outbound
Wan-S2V: Audio-Driven Cinematic Video Generation HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 823551b6-d303-469f-a713-b7cff82979e0 · inbound
UniVerse-1: Unified Audio-Video Generation via Stitching of Experts Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31fba0ea-5d3d-4267-8785-c127026cd2ba · inbound
ASTRA: Let Arbitrary Subjects Transform in Video Editing Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6e6f1077-bff7-42c6-91f1-b22c9c3940db · inbound
Understanding, Accelerating, and Improving MeanFlow Training Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3775d6b-28f1-4e2d-8dd4-07c78e054cc5 · inbound
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 81c585a6-fcc6-41ec-8429-171803c4fc2b · inbound
Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e3d810-2745-4a83-8ad8-805a95a82067 · inbound
JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39098de-37ea-4f0a-9dee-cc030938ab7f · inbound
LTX-2: Efficient Joint Audio-Visual Foundation Model Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation aeee2f62-f7d9-4b5a-93e4-43ba92da5b75 · inbound
AUHead: Realistic Emotional Talking Head Generation via Action Units Control Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b2b5bc18-be68-4dac-98bf-d3478a9c8bab · inbound
EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed9b6ca1-feca-4e93-8d37-466dc6ccce21 · inbound
LPM 1.0: Video-based Character Performance Model Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 53a221fe-6b10-4d3c-91ec-02fc5e31bc78 · inbound
Tora3: Trajectory-Guided Audio-Video Generation with Physical Coherence Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 915f1fcc-8e73-4951-b8ff-7eabc64a9178 · inbound
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3ea3c7b8-d52c-47e6-9c48-0973c96485da · inbound
Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0615e64f-6599-41bd-83d8-5bcb33534e7c · inbound
Generate Your Talking Avatar from Video Reference Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a3dc51c6-bb87-40c8-b871-35b2a91dfdc1 · inbound
EA-WM: Event-Aware Generative World Model with Structured Kinematic-to-Visual Action Fields Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 37293b50-a45c-4a27-a122-eaa91e372b17 · inbound
VidSplat: Gaussian Splatting Reconstruction with Geometry-Guided Video Diffusion Priors Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 85ee8bdd-7e3f-4f54-b055-863779596bf1 · inbound
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fe6b0a7c-2970-472b-a76c-05962dcf8ef6 · inbound
StreamChar: Long-Horizon Streaming Character Audio-Video Generation with Decoupled Orchestration Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 39e56e96-9933-4318-b211-5bfd45fd9511 · inbound
LongCat-Video-Avatar 1.5 Technical Report Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 802dd96e-2a18-45e5-94a0-861643019afd · inbound
ReFree: Towards Realistic Co-Speech Video Generation via Reward-Free RL and Multilevel Speech Guidance Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9adfd245-c7ea-4461-bb45-c772ba35572c · inbound
OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation abda8adc-4997-457f-95ab-4eb69b4bc66c · inbound
SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e55d4727-a3f8-4397-a63d-f09874dc90a8 · inbound
Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6500ddfb-d75d-4892-a023-fe235c500c26 · inbound
Vidu S1: A Real-Time Interactive Video Generation Model Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9925a27-de13-409c-9838-6f113ba0d756 · inbound
Vidu S1: A Real-Time Interactive Video Generation Model Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d80c2af-be36-458f-a467-b2f597e4d5ab · inbound
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96151756-fd25-4fe9-a567-e7ac1b8a8321 · inbound
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db1233f-e7bb-4a10-bb85-3d0ff1a05a65 · inbound
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51f7c75e-916d-4188-b3d9-db5c41a5be5a · inbound
AptAvatar: Fast and Vivid Long-Form Audio-Driven Video Generation for Production-Ready Avatars Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32ba9d0c-e283-4059-9009-9515cc4f2452 · inbound
TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbc871b1-6b14-47db-b6b1-2c0f3c82aaef · inbound
LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afd2025c-a931-4640-a65e-51dd3c8b3371 · inbound
EchoCache: Energy-Guided Cross-Modal Caching for Efficient Audio-Driven Video Generation Wan-S2V: Audio-Driven Cinematic Video Generation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.