Pith. sign in

Paper Citation Record · LEDGER

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

As of 5 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 26 inbound Pith citation observations for arXiv:2512.04677.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.04677 v6

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T18:39:43.814175Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T01:32:13.124086Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T20:00:07.549832Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 82cc5007-8824-4d43-8e61-1b0d29077801 · outbound

This paper cites Body of Her: A Preliminary Study on End-to-End Humanoid Agent.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Body of Her: A Preliminary Study on End-to-End Humanoid Agent

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.166965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.166965Z digest=sha256:898580df995dc38f5a5ce87ea627adf69be0ba0219734ee862c0153d0455d8bb

Observation 6c882b2d-70b8-48fe-9241-f11e7331f72c · outbound

This paper cites Video generation models as world simulators.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Video generation models as world simulators

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.175579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.175579Z digest=sha256:291b39782523477495bc7048dc036c72d135ec8614117760559d73bf84ce5c6d

Observation 83acffe0-41fe-4059-98eb-121e9b300650 · outbound

This paper cites Diffusion forcing: Next-token prediction meets full-sequence diffusion.Advances in Neural Information Processing Systems, 37:24081–24125, 2024.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Diffusion forcing: Next-token prediction meets full-sequence diffusion.Advances in Neural Information Processing Systems, 37:24081–24125, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.188926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.188926Z digest=sha256:346c2dc43087da13be40d499c2403eb355d12414a41189afb00958ea387282e1

Observation 0cb89d86-de96-464c-a2f5-08274b1f2d33 · outbound

This paper cites MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length MIDAS: Multimodal Interactive Digital-humAn Synthesis via Real-time Autoregressive Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.196491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.196491Z digest=sha256:469eac7796c4d2d09cfe03799a89f13473fc8a8d4423dd4f78d6d7d76da38c3e

Observation bb9c26c1-4969-4ad2-b9f9-386bf32495ba · outbound

This paper cites Out of time: automated lip sync in the wild.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Out of time: automated lip sync in the wild

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.205395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.205395Z digest=sha256:88fab0b528761e637a5cb6863d4806e70d7ec98ac74d0eec0f9bedc840131c1b

Observation ce87fdb9-d959-4db0-b7c1-2bbc735cc005 · outbound

This paper cites Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.212151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.212151Z digest=sha256:42ee856d08856a8ceacb601ef26ccc952e640183e271b7ff061b66073af9c615

Observation 6293f12b-6786-4b41-8b35-ff1be19c0e59 · outbound

This paper cites Self-Forcing++: Towards Minute-Scale High-Quality Video Generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Self-Forcing++: Towards Minute-Scale High-Quality Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.218899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.218899Z digest=sha256:0a1e455ab139c4a4d4f38d1df1ff5f0186f2c1973d2cefc488a9fc23b766ebd0

Observation b4fb6205-9393-40a4-8c36-61d27f317e2f · outbound

This paper cites Rap: Real-time audio-driven portrait animation with video diffusion transformer.arXiv preprint arXiv:2508.05115, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Rap: Real-time audio-driven portrait animation with video diffusion transformer.arXiv preprint arXiv:2508.05115, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.226981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.226981Z digest=sha256:ce10231cc431d7cede6b96776ead7d1ed7fbdf74c7ccdbb8acbba7a386ef183d

Observation f195593b-1f3c-4f72-af00-3b3ef7bebe31 · outbound

This paper cites Cosyvoice 2: Scalable streaming speech synthesis with large language models, 2024.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Cosyvoice 2: Scalable streaming speech synthesis with large language models, 2024

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.233681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.233681Z digest=sha256:6718ca73b3edb01ac452f18923f11e2299adde2f722c2eef5325e8974b945c90

Observation 930b2a88-d848-461f-aa7e-834e4d51655c · outbound

This paper cites Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Looking to Listen at the Cocktail Party: A Speaker-Independent Audio-Visual Model for Speech Separation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.244873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.244873Z digest=sha256:91cc2f2b80cd5fa3ee72f5fd7d23dfd849e9fc1125866ec3656503e391b56e7f

Observation d581221e-ad3c-4401-ad91-16328a1fc7a5 · outbound

This paper cites Phased dmd: Few-step distribution matching distillation via score matching within subintervals.arXiv preprint arXiv:2510.27684,.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Phased dmd: Few-step distribution matching distillation via score matching within subintervals.arXiv preprint arXiv:2510.27684,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.255976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.255976Z digest=sha256:356fdcfb1fae2721def632aeeba697ba029a1aecce729375dc55faf53af0ebd2

Observation d79c79e1-ba36-4e21-938d-bdf59918cd5d · outbound

This paper cites One Step Diffusion via Shortcut Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length One Step Diffusion via Shortcut Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.263893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.263893Z digest=sha256:c5f321c7a541515fbf65163d2cbb8f9eaa5fc5fa2d456b5949b9793a326e9f66

Observation 1663fec6-a4ca-4694-89e5-b22af599ca08 · outbound

This paper cites OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.281527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.281527Z digest=sha256:4dc59921463c89204fc17bdeb3224ec01e0b72a7d9ea9c32f42c0f60419761cf

Observation 021482ff-045d-404b-86df-c9c56096c618 · outbound

This paper cites Wan-s2v: Audio-driven cinematic video generation, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Wan-s2v: Audio-driven cinematic video generation, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.288890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.288890Z digest=sha256:d2dc9059b950aef06a56320706db3bf15b52531399a85051b8b7fec66eca34df

Observation 81c585a6-fcc6-41ec-8429-171803c4fc2b · outbound

This paper cites Wan-S2V: Audio-Driven Cinematic Video Generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Wan-S2V: Audio-Driven Cinematic Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.295310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.295310Z digest=sha256:7dfe50ec5e350169345400f9292799fbe8467a986c69f1c9a05b1ad21b662abf

Observation 8be3340d-0da7-46d2-818d-dcfec0a56880 · outbound

This paper cites ARIG: Autoregressive Interactive Head Generation for Real-time Conversations.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length ARIG: Autoregressive Interactive Head Generation for Real-time Conversations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.303159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.303159Z digest=sha256:d97881b7a6484f1f1ba0c3a5ba4820c024d7a492be835d4e90e8a5ce04f3b10b

Observation 8524f4d6-e8e5-4578-be4a-92bf81207454 · outbound

This paper cites LTX-Video: Realtime Video Latent Diffusion.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length LTX-Video: Realtime Video Latent Diffusion

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.318204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.318204Z digest=sha256:f93fcb0e08f515aa2b3272cc34f8e4b895b3a099ca5e2535ae453218b7512a59

Observation 63dbcf92-3228-4a37-aa08-0f00a73b0cd0 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information processing systems, 30, 2017

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.335539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.335539Z digest=sha256:b1f60e60871eca90de7a6078b2ca2331e997cb27b3fd630844617432c37e2999

Observation cfbac149-1467-4cb7-b140-b48b69acd3e4 · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.344768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.344768Z digest=sha256:ead11a0a198680149e742bdc8aac53ccf478a2bad84391b930728a9a7df08c34

Observation cc9cfed6-5d97-41d9-bf07-a26e1da6796c · outbound

This paper cites OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.351503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.351503Z digest=sha256:54287c572266af8f7535db399103eee1a1c93519382ee2a2742efb89fa633059

Observation 544bbe8d-08dc-4075-be83-c277f6817d28 · outbound

This paper cites Streamdit: Real-time streaming text-to-video generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Streamdit: Real-time streaming text-to-video generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.358078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.358078Z digest=sha256:6349ec255ff1bd3722a11fccda4c291d57012bdde4589d947451c096620299ee

Observation abc910a8-b0b3-480e-828e-57e52457be0a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.373827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.373827Z digest=sha256:4ae7b514ece83ea50369fe33d92fcd03f42f1300943d4a03c91de21d9ca4dc8b

Observation 9add171d-7f97-4bd4-8578-e5b5000b0dec · outbound

This paper cites Autoregressive image generation without vector quantization.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Autoregressive image generation without vector quantization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.393147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.393147Z digest=sha256:7b0a8bb41eba1591f69493d514ae3ff080eaaeb26e8433e76b8baa1225f021a3

Observation be87dee5-2d14-4507-a117-c2207b12a1db · outbound

This paper cites Ditto: Motion-space diffusion for controllable realtime talking head synthesis.arXiv preprint arXiv:2411.19509, 2024.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Ditto: Motion-space diffusion for controllable realtime talking head synthesis.arXiv preprint arXiv:2411.19509, 2024

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.484166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.484166Z digest=sha256:f05ac3970e6d62049bb24b4da0750474d39c741c0bbe57d47e0a1451ad617a71

Observation 2bb663be-07a6-41a9-bc80-f75257e63abe · outbound

This paper cites Flow Matching for Generative Modeling.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Flow Matching for Generative Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.581561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.581561Z digest=sha256:e13f366f1c874c2d50eab654b5b2b56c7a5a1210fbb1f6c768856f2d335f0aaa

Observation f9c82ac0-cd8a-47d3-a9b9-e4470c33fb2a · outbound

This paper cites Rolling Forcing: Autoregressive Long Video Diffusion in Real Time.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Rolling Forcing: Autoregressive Long Video Diffusion in Real Time

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.712340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.712340Z digest=sha256:484db2e30f3e5d1e1e6badc1d04955cfb166d42b424677d4f366c7f52cc6716f

Observation 4a3213b5-584f-4e40-8b81-1465a29f7d90 · outbound

This paper cites TalkingMachines: Real-Time Audio-Driven FaceTime-Style Video via Autoregressive Diffusion Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length TalkingMachines: Real-Time Audio-Driven FaceTime-Style Video via Autoregressive Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.890004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.890004Z digest=sha256:0fb3e6720bb4dcb038c32d03da6443b0f4d659c2d132b82433e7c8ad6d89a8b5

Observation 29508ad4-6461-4533-9415-ca95679d4c8d · outbound

This paper cites Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:41.980575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:41.980575Z digest=sha256:a93c56cbcfcf7d0399a7d198dbad336714be7cba9c4a000ad7eec6ce532fd773

Observation 7a651115-1861-4b43-ac1f-06cf972e5c69 · outbound

This paper cites Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Diff-Instruct++: Training One-step Text-to-image Generator Model to Align with Human Preferences

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.081824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.081824Z digest=sha256:a418944d66fb7fd8bd30e5d73b1e38ba519a42d639f6120a0186cf1a9272e12d

Observation 6df3d4ce-7ea4-41b3-8423-5b75d5b20f4f · outbound

This paper cites Diff-instruct: A universal approach for transferring knowledge from pre-trained diffusion models.Advances in Neural Information Processing Systems, 36:76525–76546,.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Diff-instruct: A universal approach for transferring knowledge from pre-trained diffusion models.Advances in Neural Information Processing Systems, 36:76525–76546,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.153832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.153832Z digest=sha256:b9db9b9150c1a6e1bfe00b95304d2589debdb59e7146ce5826812338bb412610

Observation 6a612ec4-8215-42b1-8543-4df5888466b5 · outbound

This paper cites Learning Few-Step Diffusion Models by Trajectory Distribution Matching.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Learning Few-Step Diffusion Models by Trajectory Distribution Matching

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.201403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.201403Z digest=sha256:87c7a7219bb0ad710461f9b1f4460e23a80cdffcd924ef0e63597f89f9b7864f

Observation e7d1445d-3efc-48f8-866b-590dff817efe · outbound

This paper cites MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length MirrorMe: Towards Realtime and High Fidelity Audio-Driven Halfbody Animation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.261377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.261377Z digest=sha256:09315c4b99453a143f78ffbf1b9b69b7d9c8ca4274400cadda8d9f3b0426a0e5

Observation 62f569ff-33e4-4aae-9f1b-f801de2eee1f · outbound

This paper cites Echomimicv2: Towards striking, simplified, and semi-body human animation, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Echomimicv2: Towards striking, simplified, and semi-body human animation, 2025

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.324041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.324041Z digest=sha256:ad8cc2e31064cfe5a0ad7809af4d47e393aa74e41c1e6fddb7e8738f6cb475a1

Observation 59e93b2b-a253-42dd-adca-af494b46dc89 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length A lip sync expert is all you need for speech to lip generation in the wild

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.391465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.391465Z digest=sha256:d09aca32a1f175f5f68e79870aed1caa5dc8aa6ac9d02ada26c999f68934e62d

Observation b7cc940c-c6ba-4fca-863b-4cbd4d3f1d7b · outbound

This paper cites an unresolved cited work.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.472526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.472526Z digest=sha256:018908bc2ea896792eaad332406957057b43dd5e9ec0a66e340ba6f124e16795

Observation aa6e5242-a570-4805-bed4-be3d0f9c858e · outbound

This paper cites Consistency models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Consistency models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.602085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.602085Z digest=sha256:15e59b4b1e5f4fcdf6f100bd174e746427e196b749cb7029205ca33cc0f9d4b2

Observation b1f3b75b-42d1-4bd0-b6c8-f7e5e0eb9d70 · outbound

This paper cites Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.726198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.726198Z digest=sha256:ee037302c6b1cfca854b1ff42c5f31c0866b3f35928756581eb1fb4593a27532

Observation 8ec5b113-bb5e-47f2-8bf6-efac56ea9c95 · outbound

This paper cites StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.827965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.827965Z digest=sha256:22ccd8eb9434054cdf3d43f1f616bad9d3566061ae8141d0521c54ffed709b8b

Observation 727c6acc-45c6-412a-857c-9552b1181343 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:42.932746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:42.932746Z digest=sha256:4e683ace76529001a9904f49dd2a4a733f1ff996cd3d3fe306132860e5ffb5c3

Observation 78f1af44-fc27-42c5-8430-e0ebe4f7be37 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Wan: Open and Advanced Large-Scale Video Generative Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.025308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.025308Z digest=sha256:345181da1fba9985c1d8c8560ee937ef2898a929362f1245dd854c303e5df3cf

Observation 9e9f5b62-f9b4-4460-b856-8a17fd453039 · outbound

This paper cites Fantasytalking: Realistic talking portrait generation via coherent motion synthesis.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Fantasytalking: Realistic talking portrait generation via coherent motion synthesis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.128520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.128520Z digest=sha256:8ad9b52d6637aae2454ff19f0a13fa5d840438a35e6ae92c907d6c84c59f9b57

Observation d217822f-4c81-4158-a40f-eacbfaf14068 · outbound

This paper cites Omnitalker: Real-time text- driven talking head generation with in-context audio-visual style replication.arXiv e-prints, pages arXiv–2504, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Omnitalker: Real-time text- driven talking head generation with in-context audio-visual style replication.arXiv e-prints, pages arXiv–2504, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.173394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.173394Z digest=sha256:9bec87307e57cabf0dbbdf908ed804de3ee27a2e52a32bf90873b3e5b14a7912

Observation 9d568af0-4201-43db-9b95-c408d0059081 · outbound

This paper cites Qwen-image technical report, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Qwen-image technical report, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.232969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.232969Z digest=sha256:c8b1e9697bc35eb5259ddb73963bfb75dab3e0575d0e56dc8aef3e25df0cfd4b

Observation e8cffe70-559d-4425-be2f-dadef4d00fc0 · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.295251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.295251Z digest=sha256:2b398383789071366ed5beeaf852b43e1bdea29ea8ee3a660c1d6ddb4240e476

Observation 0be907c6-0a19-4a25-b7d9-042ab3293749 · outbound

This paper cites X-streamer: Unified human world modeling with audiovisual interaction.arXiv preprint arXiv:2509.21574, 2025.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length X-streamer: Unified human world modeling with audiovisual interaction.arXiv preprint arXiv:2509.21574, 2025

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.350746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.350746Z digest=sha256:92deb3d4b5f1a5fc77b73fdbd70940c3b575365562cf8b79b9cf2f5f594e4d55

Observation 29915014-a2a3-4557-8498-8ddf07022dd1 · outbound

This paper cites LongLive: Real-time Interactive Long Video Generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length LongLive: Real-time Interactive Long Video Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.393621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.393621Z digest=sha256:e8aef5f98b3399f5ca2fb1536f6d1d7d7b55ac62e5cdbae3bb2b10118c59ba87

Observation dcfaf336-cea5-4464-abb7-4b61cee2862c · outbound

This paper cites InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.445612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.445612Z digest=sha256:8db5c68f69fbb102301d30fde3e77208728f7a602f43098b75104a436264a490

Observation 6cd75b41-b782-4efc-8903-0aabfa672d10 · outbound

This paper cites Improved distribution matching distillation for fast image synthesis.Advances in neural information processing systems, 37:47455–47487, 2024.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Improved distribution matching distillation for fast image synthesis.Advances in neural information processing systems, 37:47455–47487, 2024

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.497860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.497860Z digest=sha256:2e5bd430fba9bcc7677e6168e72b479a707cfb3471354e903a2c3769ae8890e0

Observation 6f4624c1-3451-4656-bded-75e00a55ae3d · outbound

This paper cites One-step diffu- sion with distribution matching distillation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length One-step diffu- sion with distribution matching distillation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.551659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.551659Z digest=sha256:9828e3195308987f4744405f7c4397837498524a5a9ae8739a5a4dfbd86a4de8

Observation 65cdae69-2ece-439f-8028-6d82c92bde4a · outbound

This paper cites From slow bidi- rectional to fast autoregressive video diffusion models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length From slow bidi- rectional to fast autoregressive video diffusion models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.594113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.594113Z digest=sha256:b6cdec299c0af981fa2a24ec3ed8a16d1a0408de2eb7b7a4b2ba2903de6044a2

Observation fb09a11e-2136-4d79-978a-df66026497ea · outbound

This paper cites LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.646679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.646679Z digest=sha256:a9780e1a128f489bd1893d0ff790c1ab32b116e61bfd73f20c4f8db7371a76ad

Observation 8d3fc1ae-13a4-4a63-bed2-617bbd222386 · outbound

This paper cites Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.675227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.675227Z digest=sha256:222e99806e70725b7ad6304375c1f247c946c1eb17e435dc156d42cbbd83ef47

Observation a67120ca-2a4f-48d9-a07e-fb084c0cef00 · outbound

This paper cites Teller: Real-time streaming audio-driven portrait animation with autoregressive motion generation.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Teller: Real-time streaming audio-driven portrait animation with autoregressive motion generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.715177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.715177Z digest=sha256:95cb1e4427efb2e2328453acf2500ac7834b5571b36878a0a6021183e6392634

Observation 33c64dea-f1c6-4d3d-bcff-3be202bd471d · outbound

This paper cites Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Large Scale Diffusion Distillation via Score-Regularized Continuous-Time Consistency

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.757479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.757479Z digest=sha256:6a3bdb758041a3a18a67d66fe4dc1dc4db9ae4a4579076bb3122e479d7804f50

Observation ca5c6681-a875-443a-8714-d4fd1789ff83 · outbound

This paper cites Infp: Audio-driven interactive head generation in dyadic conversations.

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length Infp: Audio-driven interactive head generation in dyadic conversations

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T18:39:43.814175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:39:43.814175Z digest=sha256:1552fd5f7a9cb0f6582fa0e91c5b2ffd020159d7410913931b4c0c4f6cc79783

Pith citing papers

Observation 99a45afe-36b5-44ca-96e6-55e0c3724223 · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-15T22:20:22.223470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:3e91152605f0bef551b8de952bfc59155397398232185c6a37bed738948812fe

Observation 1613a44a-fc5a-4f64-b077-2620caabc7e5 · inbound

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms cites this paper.

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 141

Resolution
verified exact
local_arxiv, observed 2026-05-14T01:38:36.168716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T01:35:14.878069Z digest=sha256:b0f8eea88755286460c7d3c07cf3e2db6f4563b4f09554b327efc94c4106cf04

Observation 7aa332d0-ea76-48ad-9c40-329268c0e053 · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-10T22:45:49.266963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:36:42.100191Z digest=sha256:423169068dc32bc45cfb5ca4abfc062684144a7d05028bddd28464823caa2a58

Observation ef8231ae-2667-4bbf-8895-cc2fc003d90a · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-13T09:42:23.808691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:42:23.808691Z digest=sha256:9e44e2cff2d0e6d76bdab6e6f861e5dcb094becc7cc295bef1ff975ec366166d

Observation b469a634-839d-4870-a18e-d83cecd4e917 · inbound

LPM 1.0: Video-based Character Performance Model cites this paper.

LPM 1.0: Video-based Character Performance Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:30:59.831661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:09:17.360697Z digest=sha256:57a648c4da8b00f6b03a2a2e4713ab55de5cdfc989e5db7ea69776873104b93f

Observation b2d76868-eb4e-43d0-b807-9a5534514e36 · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 300

Resolution
verified exact
local_arxiv, observed 2026-05-10T09:03:26.263831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:eff0c2f8e096654255f1fc6661811101235527884119920fc1146d4eb33360c8

Observation 395d64aa-8104-412c-9a1f-f417e39b568e · inbound

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations? cites this paper.

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations? Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-11T21:06:14.726705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:42:31.456648Z digest=sha256:4c086bd34d2d4b8bf25a377c02e6d57956c313bd5feace3a0baa239627219d51

Observation d385db71-4ea4-4b20-9de2-a3e220d19151 · inbound

Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models cites this paper.

Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-12T06:51:28.014977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T03:56:25.472887Z digest=sha256:e3463f17678c7d07988c69fe4a0e1f5702f2bce0b4ce172198080b0168885a22

Observation c330c1f8-ecf7-4b1d-a4eb-ebdf59e968f9 · inbound

CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives cites this paper.

CausalCine: Real-Time Autoregressive Generation for Multi-Shot Video Narratives Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:42:21.265641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T05:38:21.587653Z digest=sha256:6f8a04305cd52d802554b3f6ed1028e5b884813f783d615b9b3e8b8b975b6196

Observation 8fe64862-0510-4907-b932-acbf5047c503 · inbound

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation cites this paper.

Causal Forcing++: Scalable Few-Step Autoregressive Diffusion Distillation for Real-Time Interactive Video Generation Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:05:04.356614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T20:59:34.847496Z digest=sha256:a05594b9d75bec7c024aa6924193087918f98611a699415150b55bd5c58bf3b3

Observation 6fd0936e-c7c2-437f-87da-36a07ebbdeb3 · inbound

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization cites this paper.

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-20T20:13:43.552039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T20:11:33.277349Z digest=sha256:91989300fc4e843d3d7ccfd2784a106e01ed496fc9c9121285d5dae3b6e42026

Observation b7229994-1b38-4141-835f-f6c850336a4f · inbound

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization cites this paper.

FashionChameleon: Towards Real-Time and Interactive Human-Garment Video Customization Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-30T19:25:00.595707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T19:23:07.044659Z digest=sha256:610fd0695e878f5afa7287ed5c96e8542802564da301db187680377ab8202823

Observation 7040923c-9aff-48c2-ad82-54f78f4ce2fe · inbound

StreamChar: Long-Horizon Streaming Character Audio-Video Generation with Decoupled Orchestration cites this paper.

StreamChar: Long-Horizon Streaming Character Audio-Video Generation with Decoupled Orchestration Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-06-29T22:24:00.828792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:15:17.804936Z digest=sha256:db063bc9301f1b9c47c7bb52dd111f20d059915d018c840a6761bb1b2dcdbd14

Observation a3054365-5e85-48ab-a15d-d5981e0af8a0 · inbound

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models cites this paper.

minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:23:14.768411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T08:22:48.074724Z digest=sha256:14967ade8955389526dafdb01cde43bd39c49605a35cb6965fdc1c50fdada425

Observation 968f2d3e-6397-4882-9859-6b793af4a90a · inbound

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization cites this paper.

Lip Forcing: Few-Step Autoregressive Diffusion for Real-time Lip Synchronization Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-03T04:47:38.045330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T13:40:14.506288Z digest=sha256:7887aff11e67492f42b1d37d32610b91565071d2a5e2ad4c241bfd0084e5535c

Observation d644cd62-1d19-40a8-b50e-07e65f48a20f · inbound

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model cites this paper.

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-07-03T20:48:55.732426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T01:09:02.214590Z digest=sha256:dcc655879490fe178dd8ff0798f06c335d7f0ee8a1354c7ebedafe83a4daabcb

Observation f2c5dc92-729e-44a0-bb40-35e7883ccabd · inbound

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars cites this paper.

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T10:09:44.558809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T09:09:06.925645Z digest=sha256:3945f5cc4e4839df201c0f004b6eefe883cedb97ca585ec23d293fbb963ce194

Observation 4a6eda4c-79e3-4002-86ef-bd3402d610ab · inbound

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars cites this paper.

InteractiveAvatar: Real-Time Streaming Video Generation for Consistent and Intent-Aware Avatars Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T07:05:29.126354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:00:53.496569Z digest=sha256:1d06fcc2e154191ca67a5b9e3ec98f310a57413569586f504f092839e064837b

Observation cabe0c64-4d48-4b5c-a328-0f3353363351 · inbound

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models cites this paper.

Causal-rCM: A Unified Teacher-Forcing and Self-Forcing Open Recipe for Autoregressive Diffusion Distillation in Streaming Video Generation and Interactive World Models Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-04T20:00:07.552738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T20:57:30.765802Z digest=sha256:555d74051e70117df96c96b5a139e1c501b32e99072fee8b1caab00b8b67c6ca

Observation 3076cf5c-aefb-4d2c-97f9-974b1dfc8c0e · inbound

Towards Memory-Efficient Autoregressive Video Generation via Instance-Specific Parametric Absorption cites this paper.

Towards Memory-Efficient Autoregressive Video Generation via Instance-Specific Parametric Absorption Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T14:17:02.583722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T14:12:29.528065Z digest=sha256:69e9313402a1acfaa9d00fbeb91d6de647f3422d18d3c2c796f9dbc228bae8cc

Observation 2b5db40c-d9d9-4e59-be53-d991f80c4163 · inbound

Vidu S1: A Real-Time Interactive Video Generation Model cites this paper.

Vidu S1: A Real-Time Interactive Video Generation Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T04:46:57.581402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T04:46:57.581402Z digest=sha256:3793f2a645b126845932b62e7657303487d44b4f1254805ab1550fa4c84e509f

Observation 40026c4a-3386-4c42-9dce-a3c30f010f97 · inbound

Vidu S1: A Real-Time Interactive Video Generation Model cites this paper.

Vidu S1: A Real-Time Interactive Video Generation Model Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T08:56:31.566881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:56:31.566881Z digest=sha256:c3f8dbcc27bafb01060328112c9a9e5348f9b4d2ca606a0492e9c9ad56da9bf4

Observation d64ade22-3c3d-4a9b-8c53-dd76bc0c4b5d · inbound

OmniMate: Open-Ended Real-Time Streaming Audio-Visual Generation for Interactive Avatars cites this paper.

OmniMate: Open-Ended Real-Time Streaming Audio-Visual Generation for Interactive Avatars Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T03:52:55.366559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:52:55.366559Z digest=sha256:ea97b2b49c88d65dbfbb693b50f250d19c76f1e68552e3c492cdf4d2e3703c4d

Observation 8cbe9111-a429-44bd-9408-8e140c0c33b2 · inbound

TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation cites this paper.

TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-31T17:08:11.655887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T17:08:11.655887Z digest=sha256:742d49bfb5ffbed483c842b7f243c2d4e1536a7ae3588d6de154ba11b2885531

Observation 00153b5a-8d9e-4c80-8ac8-93c7c3184ff8 · inbound

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation cites this paper.

Rethinking Classifier-Free Guidance in On-Policy Diffusion Distillation Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-07-31T06:35:52.560992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:35:52.560992Z digest=sha256:2817956029bb789d57d417bacfa4d76e9e12c7ed9aaebe41b29e677da10a0160

Observation 1dd65f5d-9171-4ebd-b592-5ac3970ae992 · inbound

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation cites this paper.

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-04T01:32:13.124086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:32:13.124086Z digest=sha256:b935ec161a1db71e4fdf2dd2e7a1eb66b287f8f1c3e3f5cf8c7aab402dcae31c