Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:41:59.683979Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 0 inbound Pith citation observations for arXiv:2606.01620.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T15:41:59.683979Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
78 of 78 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c813f42c-f684-4bae-b51c-48dfe9a889d5 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ca6265-f657-48e3-a336-382951465e1c · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4c9f3cf7-807c-4417-a283-666403fc1a0a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs V oice puppetry
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c8bda84-1a96-47c1-935f-cb4ea9608c49 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Video rewrite: Driving visual speech with audio
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5739b4dc-38b7-45ed-b468-484c6a477c62 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Diffusion forcing: Next-token prediction meets full-sequence diffu- sion.Advances in Neural Information Processing Systems, 37:24081–24125, 2024
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf2927df-d842-44d5-ab01-d491843454d1 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d0c54751-da2e-4982-928b-52405f182036 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Dc-videogen: Efficient video gen- eration with deep compression video autoencoder
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 84883b82-5ab2-4acc-9875-908523c39f90 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Lip movements generation at a glance
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8fb1432-b4cb-4bfd-9ad5-3c106087b9c5 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Echomimic: Lifelike audio-driven portrait animations through editable landmark conditions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3bff521-6d9e-447b-93e2-9a13aa0a305f · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Out of time: auto- mated lip sync in the wild
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abad4ece-1d32-4592-a098-be24efa0a62a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs VoxCeleb2: Deep Speaker Recognition
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd568a8e-1ed5-45e7-b378-c40b135ec2fe · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Hallo2: Long-duration and high-resolution audio-driven portrait im- age animation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d182d93f-391d-4f3c-9894-f5b9293c93d7 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68a15157-24ca-439d-ab10-cf9f7192b1a6 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs EMOPortraits: Emotion-enhanced Multimodal One-shot Head Avatars
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 78929906-de10-4040-bc24-d20d1778260e · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa4bb31-951f-4461-b8bc-bbda93ad11fe · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Introducing gemini 2.5 flash image: Our state-of-the-art image model.https://developers
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b3e1e87-c56e-44ce-a5b9-47ae08da7fc7 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 128d84e6-e17b-41dd-8ebd-9da7e4358dd6 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Space: Speech-driven portrait an- imation with controllable expression
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bfd662d-e545-45f4-bd5b-b2c13df345b6 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs LTX-Video: Realtime Video Latent Diffusion
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c7086045-68a1-4c48-a328-e741cc56d020 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Classifier-Free Diffusion Guidance
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 404d936d-962b-45dd-bf0b-9a73af6c4195 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Denoising dif- fusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32d6b69a-ae24-4da6-bd77-39fd3d3a2434 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 348273fd-297f-4451-865d-48fcf9465b53 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Eamm: One-shot emotional talking face via audio-based emotion-aware motion model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c439a47e-ae37-440c-992f-82413fc84d00 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Sonic: Shifting focus to global audio perception in portrait anima- tion
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7af8d6f1-de50-408a-88c2-75fbd8775aab · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Auto-encoding vari- ational bayes, 2013
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd95b62a-58e9-4853-94f9-a9075aad1e5a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Tokenmotion: Decoupled motion control via token disentanglement for human-centric video generation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77dd19b0-02c9-4605-a2ad-aed09b59b4e4 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Expressive talking head generation with granular audio-visual control
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f06aec-259c-4c30-9bd3-b12057a11376 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 28fe0d3c-60e9-4a85-8542-8a889451a53a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Diffusion adversarial post-training for one-step video generation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d20be360-94c1-42cf-9542-e8a98fa6d7ef · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Autoregressive adversarial post- training for real-time interactive video generation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 42eb07bb-d5c3-464e-a546-3a0bc46992d7 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Flow Matching for Generative Modeling
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2578d23f-a9fa-472f-9cc1-2a4f51157076 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8824b76c-9096-4e16-acbf-911c9d25e250 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs TalkingMachines: Real-Time Audio-Driven FaceTime-Style Video via Autoregressive Diffusion Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 63907d27-5bd1-4706-acdd-6f4e2e0b7191 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Scalable diffusion models with transformers
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 837f0016-0877-4393-a5be-724c0ca2cb0b · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Open-Sora 2.0: Training a Commercial-Level Video Generation Model in $200k
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 080d370c-711e-4619-8f20-38259de1b6f4 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs A lip sync expert is all you need for speech to lip generation in the wild
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d79fc0-424d-4c60-9975-3ef5356e3887 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ca902b09-28de-436b-8798-6240138994d0 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs High-resolution image synthesis with latent diffusion models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3f0d381-da78-4aa0-ad2b-15a1ccd747ec · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Fast high- resolution image synthesis with latent adversarial diffusion distillation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d598ce8a-f2a1-459b-b5bc-5b82075bb1c6 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Adversarial diffusion distillation
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c11c598-c407-470b-837c-4e473c7f0831 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Score-Based Generative Modeling through Stochastic Differential Equations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 382db98f-2f1e-4f0e-82f9-a28b56fe37d3 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Diffused heads: Diffusion models beat gans on talking-face genera- tion
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed3414aa-3d56-43ed-86b7-623ea378a884 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Synthesizing obama: learn- ing lip sync from audio.ACM Transactions on Graphics, 36 (4):1–13, 2017
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1207205d-4708-4eba-b7f2-6443f6133cc3 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs MAGI-1: Autoregressive Video Generation at Scale
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6c77681b-e29f-4049-af64-1e9cda9efae0 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e098ef0b-d711-436a-85db-e978b75d2a8b · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Reducio! generating 1k video within 16 seconds using extremely compressed mo- tion latents
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ad01a55-9503-45e5-a82e-a7492f7f1d9e · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Fvd: A new metric for video generation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e25ec6e8-9d5f-40f5-b002-445fa2f40815 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Conditional image genera- tion with pixelcnn decoders.Advances in neural information processing systems, 29, 2016
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c3dd92-4805-48e8-88e7-cb56beb32e2a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Pixel recurrent neural networks
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c06e113e-91ef-48c1-b6f4-fef71baf7dec · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Wan: Open and Advanced Large-Scale Video Generative Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 19201b67-4902-4a91-93fe-d0e67f4cf536 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Progressive disentangled representation 10 learning for fine-grained controllable talking head synthesis
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab4fd07-4d3b-4af9-9191-04be2ccc08eb · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Echoshot: Multi-shot portrait video generation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 274846f4-57e9-4428-ae7d-daf5365ca4d3 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Fanta- sytalking: Realistic talking portrait generation via coherent motion synthesis
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98847c52-c98e-4f91-a1a1-970789955024 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Audio2head: Audio-driven one-shot talking-head gener- ation with natural head motion
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de4bb315-8d87-4484-ac21-c4228792c5b2 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs One-shot free-view neural talking-head synthesis for video conferenc- ing
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 679e79a2-4c15-4b52-8c6a-2414065c442f · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0019adac-4292-42ae-8420-c522e548afc7 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs A learning algorithm for continually running fully recurrent neural networks.Neu- ral computation, 1(2):270–280, 1989
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a0bf6b9-050b-4aae-9500-3c2f8ad4609f · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 56ac3d45-0d76-49e0-af35-1bfce50c0b49 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Vasa-1: Lifelike audio-driven talking faces generated in real time.Advances in Neural Information Pro- cessing Systems, 37:660–684, 2024
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d6057f-fa6e-4fce-81af-0455903a8b65 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs VideoGPT: Video Generation using VQ-VAE and Transformers
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c6b6c8bd-bace-48d2-b730-5f653fedd40a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b1aca6fa-de05-425d-afa0-29fcaae31a06 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6893e448-5f43-4a37-a0aa-e5f020447700 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Styleheat: One-shot high-resolution editable talking face generation via pre-trained stylegan
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2fe0017-c206-4946-95fc-d2ba933d0ca8 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Im- proved distribution matching distillation for fast image syn- thesis
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92b13d9e-0fe2-4156-af8f-6e531913be1a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs One-step diffusion with distribution matching distillation
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6e4521e-524d-46cf-8bbf-a039197b53eb · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs One-step diffusion with distribution matching distillation
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ba160af-7d76-4209-ba35-a4e79a86cff1 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs From slow bidirectional to fast autoregressive video diffusion mod- els
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a6c1562-5294-4ff8-9492-01538a1ad6dc · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs An image is worth 32 tokens for reconstruction and generation.Advances in Neural Information Processing Systems, 37:128940– 128966, 2024
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26b29c8c-6b41-4f2b-aa04-8677dde9e0c0 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e97a59e1-a265-4adf-a3b4-1f20ce1bd1fc · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Talking head generation with probabilistic audio-to-visual diffusion priors
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a44d403c-9d15-47d6-a899-e0b09b9fea29 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Identity- preserving text-to-video generation by frequency decompo- sition
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fce1d9d8-c618-4f57-84b6-70d7487ca5c9 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Root mean square layer nor- malization
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0d344b0-bbb0-4bf3-97e8-3b1cdd83a99c · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs The unreasonable effectiveness of deep features as a perceptual metric
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d60297-8308-45c1-b5be-8c109b7f30a0 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Sadtalker: Learning realistic 3d motion coefficients for stylized audio- driven single image talking face animation
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c498c47c-c7e2-4d2e-8720-73b1c5d9b11a · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Flow-guided one-shot talking face generation with a high- resolution audio-visual dataset
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42f90887-6782-47b2-92ca-d20516d65fbf · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs 11 Taming teacher forcing for masked autoregressive video gen- eration
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cef9d99-0165-4744-a646-493882917e41 · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Pose-controllable talking face generation by implicitly modularized audio-visual rep- resentation
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f1b0e0d-86c6-43cf-bbb9-693e66a7be1d · outbound
Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs split-first
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.