Pith. sign in

Paper Citation Record · LEDGER

AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 86 inbound Pith citation observations for arXiv:2403.17694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.17694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 86 of 86 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:36:09.310496Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T02:16:26.575398Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 01855afa-757d-419b-9cb5-40e207cc27d3 · inbound

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation cites this paper.

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:03:12.558326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T16:59:57.727035Z digest=sha256:09bd1940006eb004944219ca4f3cc2de6022e8f40fd7f119ac3518132c4ce8bd

Observation 2c67bbdc-c659-4e02-acc2-5376d5dc115e · inbound

Sonic: Shifting Focus to Global Audio Perception in Portrait Animation cites this paper.

Sonic: Shifting Focus to Global Audio Perception in Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T13:17:46.724519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:17:46.724519Z digest=sha256:edf21c6e191da31af1d32237fddbbad86623a0373d66c3820157b8322c37ee7d

Observation 2d6248d8-5f62-41e7-bbb3-5717e5372a63 · inbound

EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion cites this paper.

EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T14:20:34.733242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:20:34.733242Z digest=sha256:a36877ad30087e9c05e9e7d0eef884ba5d7f5b257f9254cd37c990c32f17cc4f

Observation a0475f1b-34d4-4481-a77b-500538431ca7 · inbound

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation cites this paper.

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:23:15.361627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T17:19:51.411937Z digest=sha256:d724c2a87a496f675c6048192150bb6d94a5d8e2fe3e3e1fae238773c2a23f75

Observation 9f667010-1028-4d2b-baf8-657cc0abd443 · inbound

Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook cites this paper.

Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 179

Resolution
unresolved
no resolver link, observed 2026-08-12T10:08:48.971441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:08:48.971441Z digest=sha256:f5004bada6ca72e7a6f034a1588286190a56a698ae3221714e6db38811490f77

Observation 128d418f-ce0b-4bfd-9d67-6b07a75420ec · inbound

Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation cites this paper.

Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T05:08:08.994747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:08:08.994747Z digest=sha256:1e9439694f2a6c5ab4b8c9ea4fe34cc2870a63eb162a40f9556ce2f61ca0d5bd

Observation 911a26f4-21e6-41da-bdca-5f400b170378 · inbound

Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer cites this paper.

Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T05:06:43.884561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:06:43.884561Z digest=sha256:106f1558b62289b9f70368ade4234b29b3c9c00b686f443a5ece60721646fbf8

Observation 9834dd9a-38ab-4947-b585-c9801ceea50b · inbound

SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model cites this paper.

SINGER: Vivid Audio-driven Singing Video Generation with Multi-scale Spectral Diffusion Model AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T22:28:08.047726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:28:08.047726Z digest=sha256:8415acedf909d6c5f85652a65c0d2dd54f233674885bffe46ff0019bc40bd7f4

Observation 7c1e4e87-f78f-431d-85fe-76ef3d93e26e · inbound

Sprite Sheet Diffusion: Generate Game Character for Animation cites this paper.

Sprite Sheet Diffusion: Generate Game Character for Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T22:15:12.054746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:15:12.054746Z digest=sha256:0a4db95038f43cd8729b14ac8f129edf9a38655efddd8dd8358ab42414e9f398

Observation 4c6f8ae8-a1dc-4bed-ab43-503ab91a9a37 · inbound

IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation cites this paper.

IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T21:56:37.899547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:56:37.899547Z digest=sha256:1b8426926f6dc2ab55442b173ec37725a8b82d2b5298a0122a01f50ff9ee4901

Observation f4610538-afd6-4019-b79b-bac7abee8b26 · inbound

MEMO: Memory-Guided Diffusion for Expressive Talking Video Generation cites this paper.

MEMO: Memory-Guided Diffusion for Expressive Talking Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T21:28:26.512290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:28:26.512290Z digest=sha256:9fc694b4f5f957b7ce2519408549f8ff8963b818e3d7fbabf41f6aab754bbbef

Observation 1e5564bf-b52d-4aef-af79-4ee10ba190e5 · inbound

Real-time One-Step Diffusion-based Expressive Portrait Videos Generation cites this paper.

Real-time One-Step Diffusion-based Expressive Portrait Videos Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T13:09:59.702072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:09:59.702072Z digest=sha256:364376f463fb7d948a7669bbe96bb2e2d685e9052c0fcba9d57f58661cc53eb3

Observation c544c2b8-7ca1-46a2-882a-bc62f626d36e · inbound

RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global Complementation cites this paper.

RealisID: Scale-Robust and Fine-Controllable Identity Customization via Local and Global Complementation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:19:35.174108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:19:35.174108Z digest=sha256:1bd4b16bb86c29a2f4f7a3098297cf8124c7544f03a5080c11c7c31431228558

Observation c882135a-a104-48be-af75-b25eff36de2a · inbound

UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control cites this paper.

UniAvatar: Taming Lifelike Audio-Driven Talking Head Generation with Comprehensive Motion and Lighting Control AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T01:04:02.573909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T01:04:02.573909Z digest=sha256:eb333eae76addd5a56b2fe14a161e74877f8e8badbf3368d7df7c594672a1f70

Observation 040ab03c-07b8-45d8-946e-ee90a48bfd35 · inbound

JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing cites this paper.

JoyGen: Audio-Driven 3D Depth-Aware Talking-Face Video Editing AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:23:14.855424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:23:14.855424Z digest=sha256:de42fbe5e048cb3ad7a541ad0ba5264dc58d72a0386fd1f26a62653eefa400b7

Observation de69b84c-a961-47ac-93f3-bca3542fc65a · inbound

MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation cites this paper.

MoEE: Mixture of Emotion Experts for Audio-Driven Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T22:23:50.736169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:23:50.736169Z digest=sha256:13cfd743ce6717b08405b3881e8e6bf9c873be899ca7b517c7aad70bd2dbce3e

Observation e7c0954f-f1e7-4092-9389-10f6f9747d3e · inbound

Identity-Preserving Video Dubbing Using Motion Warping cites this paper.

Identity-Preserving Video Dubbing Using Motion Warping AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T21:33:20.039736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:33:20.039736Z digest=sha256:4f7f1cf97c0b9ee627c512576ef52ca87c35ee8ce3d29eab93bbce4f9d7493b2

Observation d46ca007-1ce9-453e-bca6-8a40456cc75a · inbound

Joint Learning of Depth and Appearance for Portrait Image Animation cites this paper.

Joint Learning of Depth and Appearance for Portrait Image Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-10T20:25:06.208245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:25:06.208245Z digest=sha256:91b7adc7490b3623123829c46414ac5ddfb6093780165f7fa61576cf9a92833f

Observation d42f0192-0043-4f39-b319-72e24c94825d · inbound

X-Dyna: Expressive Dynamic Human Image Animation cites this paper.

X-Dyna: Expressive Dynamic Human Image Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:52.234300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:52.234300Z digest=sha256:9cbfe2b3a15e6f64916c6d66d03d2d41d2ef0d1864733699917484cd73939881

Observation 88ec1c50-ff40-442b-8334-1dada8c4bcc6 · inbound

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model cites this paper.

Long-Term TalkingFace Generation via Motion-Prior Conditional Diffusion Model AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-07T21:11:25.341871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T21:11:25.341871Z digest=sha256:e2c7815471419d4a8a170b8d422407ac47214b0ca3d42539dcc4fbc25849f29f

Observation 3ace033b-8295-4d14-82bc-f364369fe8b6 · inbound

FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image cites this paper.

FaceCraft4D: Animated 3D Facial Avatar Generation from a Single Image AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:36:09.310496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:36:09.310496Z digest=sha256:abad5ed12b59b46bda085c6fcb2150c12cee761f7b0dd9406c087f9329a5a2b0

Observation 47f42cec-7236-4df0-a678-b351ec6dbc86 · inbound

Disentangle Identity, Cooperate Emotion: Correlation-Aware Emotional Talking Portrait Generation cites this paper.

Disentangle Identity, Cooperate Emotion: Correlation-Aware Emotional Talking Portrait Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:28:21.591558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:28:21.591558Z digest=sha256:832b39139686b1f38634f25376d3441d9b821746a0573a3dff93443f9419e4c0

Observation 29a0538a-5a70-427e-be9b-db4061972ed7 · inbound

A Unit Enhancement and Guidance Framework for Audio-Driven Avatar Video Generation cites this paper.

A Unit Enhancement and Guidance Framework for Audio-Driven Avatar Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T23:53:10.673843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:53:10.673843Z digest=sha256:1d17d80896fee061e7c8a2290f184a4c7a536bf062ae35d9f8a6df87d72cec24

Observation 99abb855-2aa8-41d8-bcf9-48c9f97b5fa2 · inbound

Test-Time Augmentation for Pose-invariant Face Recognition cites this paper.

Test-Time Augmentation for Pose-invariant Face Recognition AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T21:40:29.828545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:40:29.828545Z digest=sha256:ca74d0a86196bce3889e4ce2c1ac47100275a6bbbb907c47f1f4f5d9dd752115

Observation 1f0bcfd1-d675-4f87-ae55-6c336198b589 · inbound

Exploring Timeline Control for Facial Motion Generation cites this paper.

Exploring Timeline Control for Facial Motion Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:52.852981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:52.852981Z digest=sha256:7be42f1953ee54127dfea8a0a2e8e00ca80a3e42d8eca95946426ce799ef07dd

Observation 0d178024-8051-4ec1-a070-bf1c516af756 · inbound

FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing cites this paper.

FaceEditTalker: Controllable Talking Head Generation with Facial Attribute Editing AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:46.497816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:46.497816Z digest=sha256:c0bc0f9a5201ae46219e84e48394d104ef2fe8e65ab7312091bbf95ce66716bd

Observation b146ccef-7509-4d14-9277-f7205658eaec · inbound

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation cites this paper.

Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:34.059612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:34.059612Z digest=sha256:c18da040eaf45a8e18566f4421f06ba5b8af4f0b747cd45b0a87edc547490970

Observation 67cedca9-14dc-44f2-b242-aff70eb235b5 · inbound

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers cites this paper.

SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.767656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.767656Z digest=sha256:1d5b84b39ea1fdc7c37deae781931b8b7cce4bc0f910c25d14b96c68ce7b9220

Observation 9ac92218-4b33-48cd-bef9-b7b07723083f · inbound

Speaking images. A novel framework for the automated self-description of artworks cites this paper.

Speaking images. A novel framework for the automated self-description of artworks AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T13:17:56.065494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:17:56.065494Z digest=sha256:82fbcb1ead6fa7eb9197fb51bd7462485fa5bed7d86b5ce8c217b375d19c368a

Observation 8dce6ce4-191d-456a-94f1-149be2ffcbab · inbound

LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models cites this paper.

LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:24.450551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:18:24.450551Z digest=sha256:cbdc4cfb1eaaa049ddc54e04278c32d66ea54461964e17caaccb41892a09d5ea

Observation c3a1929b-d64c-4a7e-b2a7-33342cbf61a2 · inbound

Audio-Sync Video Generation with Multi-Stream Temporal Control cites this paper.

Audio-Sync Video Generation with Multi-Stream Temporal Control AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:24:30.408110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:24:30.408110Z digest=sha256:e8c1345832c844cd88bd49790794f91d12437905d6f21fc76f500d3d88c8e4d3

Observation e7768a50-d62b-4394-bd18-c1521d85f809 · inbound

Controllable and Expressive One-Shot Video Head Swapping cites this paper.

Controllable and Expressive One-Shot Video Head Swapping AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T19:23:24.633551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:23:24.633551Z digest=sha256:b378a6d8dc7be64105bc09b5953472e6c1b1a9c07a150e39becf4df368bb4284

Observation 05771ebd-78e4-4171-80fe-445cf21b054b · inbound

OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation cites this paper.

OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T18:47:49.019664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:47:49.019664Z digest=sha256:e4f97c14c1e54de141c847b81ab86eface86d1442e54d746419b1d66f6a2404e

Observation 042642b1-40a0-4eb1-a81b-f1ab250bac94 · inbound

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching cites this paper.

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:50:51.067493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T00:46:39.196042Z digest=sha256:8ab48762875068be9b6597d63393cd8edad93a521e4346637c9a7193a291edc7

Observation b3eff734-9988-4b39-90e7-632f17a4d897 · inbound

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases cites this paper.

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T20:58:30.336835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:58:30.336835Z digest=sha256:ddbf0089df5d0aead280315ac4153cba2fc381987cf13334313ca5eb12890b68

Observation 4654cc85-ca5c-4d2b-8181-5c1c68606216 · inbound

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation cites this paper.

MoDiT: Learning Highly Consistent 3D Motion Coefficients with Diffusion Transformer for Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T19:38:20.556209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:38:20.556209Z digest=sha256:3d0a63d2d037e60327cb5375550a07513cbbbbbb3bc875fbba34f3d04ed39037

Observation 2b7a5166-c530-4d93-928f-bb09d9a70d22 · inbound

Democratizing High-Fidelity Co-Speech Gesture Video Generation cites this paper.

Democratizing High-Fidelity Co-Speech Gesture Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:00:03.247582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:00:03.247582Z digest=sha256:05d605011aac7adfaace7d8d14bb5101dde4ec6d7d40b15f34b746d3f55494e1

Observation 8e5fc1b8-01d5-4dad-9028-01b91495ccec · inbound

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation cites this paper.

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T16:44:38.865416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:44:38.865416Z digest=sha256:b4e728beef9b5280ed904224db3b329b3da3d503d32076d175132e9ba4c0f937

Observation 451ecb1f-5204-42bb-bf13-a7b8433af352 · inbound

MagicAnime: A Hierarchically Annotated, Multimodal and Multitasking Dataset with Benchmarks for Cartoon Animation Generation cites this paper.

MagicAnime: A Hierarchically Annotated, Multimodal and Multitasking Dataset with Benchmarks for Cartoon Animation Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:37:09.367614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:37:09.367614Z digest=sha256:1f26b0369af343ce26d052cdd754a3c11a6404ac65b965c63b0339a2f20e32f6

Observation 57b2f6d9-df91-483c-b7ec-79d5c4258cc8 · inbound

X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention cites this paper.

X-NeMo: Expressive Neural Motion Reenactment via Disentangled Latent Attention AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:44.884845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:03:44.884845Z digest=sha256:5c22774536891809f3af19c966f823a62de14cb606de4f34ba9deb45acffcc6c

Observation 247b5a6b-0495-44f0-bebb-21bf2c18ce6e · inbound

Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering cites this paper.

Text2Lip: Progressive Lip-Synced Talking Face Generation from Text via Viseme-Guided Rendering AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T05:05:49.133791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:05:49.133791Z digest=sha256:61003f3a39f6498319c1fbf97158e653ed91e14ea70af01e7e9d78a59792f4a0

Observation f0271651-de71-42ed-baa9-1453bd3231fa · inbound

PoseGuard: Pose-Guided Generation with Safety Guardrails cites this paper.

PoseGuard: Pose-Guided Generation with Safety Guardrails AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T05:02:04.490234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:02:04.490234Z digest=sha256:a65988d3de0f7820da205b1cb4f415716b8f37a602facf03e9a2ce6390338e8d

Observation 499289a0-6f66-421d-af3e-61575593d220 · inbound

DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation cites this paper.

DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T12:43:23.684599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:43:23.684599Z digest=sha256:76f38963a44c23f1bdf8d045f69d7a6855007b9525da0d58f69bf13fe74f9ed5

Observation b0f0a0fb-6853-4d42-82b3-bdd80fcd9c44 · inbound

LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation cites this paper.

LaVieID: Local Autoregressive Diffusion Transformers for Identity-Preserving Video Creation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T22:04:35.670348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:04:35.670348Z digest=sha256:323705cf641928f05e1262faac28c7fa02a529f667820b66fc315f86d87ad562

Observation 3eb16f06-c3a9-4754-9280-3cd1f15c2b3c · inbound

StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation cites this paper.

StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T17:42:41.497013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:42:41.497013Z digest=sha256:82def5ffeb48c4cfb27296a09f2f05b170aa3048e9d871dc2b6151e41a89be15

Observation e80c8543-d238-47bf-8d0b-1784c00321a0 · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:49.448579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:49.448579Z digest=sha256:e68badcfcd54ab70971dc5e962b5beb127462d679cf321c6cc918d716971caee

Observation 4023f325-6310-4ae0-81b6-c88c7ac8add9 · inbound

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis cites this paper.

EDTalk++: Full Disentanglement for Controllable Talking Head Synthesis AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T19:07:36.310450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:07:36.310450Z digest=sha256:da99478834083c140a0c1e4954968bbcbf88b69036449bfa28fe454c78ad5517

Observation 9783151b-e320-4907-8b56-0523fa130e00 · inbound

TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis cites this paper.

TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:18:09.282489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:18:09.282489Z digest=sha256:414e38ae726bee02b22554ce0ba00c8375744698685a4aa54f7e29a747ee1b5c

Observation 579ea692-c4df-4f52-9229-a7b404e6c94e · inbound

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing cites this paper.

InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T18:50:16.290012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:50:16.290012Z digest=sha256:63474bc84958e9612eff25382f56a6ab235c9b2775d78b2edabc4c811cf21565

Observation a903f0a0-1513-4a28-ad24-aca133cede0d · inbound

OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation cites this paper.

OmniHuman-1.5: Instilling an Active Mind in Avatars via Cognitive Simulation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T16:57:29.778073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:57:29.778073Z digest=sha256:a316a08ea264ca6092ecefbe26017f4d0b3a32abed13bf1021fff448d69f6d19

Observation 2728a245-1ee2-4b8b-9c40-ab9ca16ade0a · inbound

InfinityHuman: Towards Long-Term Audio-Driven Human cites this paper.

InfinityHuman: Towards Long-Term Audio-Driven Human AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.244894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.244894Z digest=sha256:aacececff1407ff8c759b743f7508f471c795d8bb2badac81ee97c0205d73991

Observation d021f543-f425-431b-8fc4-9e227d662046 · inbound

Human Motion Video Generation: A Survey cites this paper.

Human Motion Video Generation: A Survey AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:57.303214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:57.303214Z digest=sha256:4816838bd414f66243f8493abe195b359fd2e3ff91a2baf4395e68bfab67c2bb

Observation 3b8687d5-0175-49e5-9919-cbe648d8ec76 · inbound

MFFI: Multi-Dimensional Face Forgery Image Dataset for Real-World Scenarios cites this paper.

MFFI: Multi-Dimensional Face Forgery Image Dataset for Real-World Scenarios AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T16:27:56.782867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:27:56.782867Z digest=sha256:f3fd13fcb39d91c1d9681df8257f99a3247494e4304349c0fb452a9da1d4721b

Observation e89d7844-db38-42da-8e1b-1700a9abf86a · inbound

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling cites this paper.

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:42:43.813294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T16:42:25.803856Z digest=sha256:5191d893a80dd9d2eb4be4655c93ca73b9a2f106cb7fc8e81243fd7890b811ed

Observation ff743ba2-f1ba-4f2e-a92c-58fd02cb41be · inbound

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits cites this paper.

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T16:30:39.321612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:30:39.321612Z digest=sha256:d3388ab0868cf593d0dda05d07d61718d34143e87b4e9ebfa850241828e3b87b

Observation be3101ce-5b9a-497a-a0f2-5d72983df5a3 · inbound

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body cites this paper.

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 113

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:08:36.327070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T22:04:07.403410Z digest=sha256:1e7eeb1deee9b40e77f24f1200e13a4d53327b2c034d613f6cfec2686f73c531

Observation 01552f3e-7d80-4027-934f-60e289d09be6 · inbound

Instant Expressive Gaussian Head Avatars at Over 100 FPS cites this paper.

Instant Expressive Gaussian Head Avatars at Over 100 FPS AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-03T15:28:59.211811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:28:59.211811Z digest=sha256:9e854e87e1eeab91cea337df9fbf6b0687ddaf90a61078fc0aa7026d6c42c706

Observation 75c95618-e92d-4942-898b-04cdcaf3db5c · inbound

UIKA: Fast Universal Head Avatar from Pose-Free Images cites this paper.

UIKA: Fast Universal Head Avatar from Pose-Free Images AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-22T12:04:51.573998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T12:02:27.828968Z digest=sha256:212fb725e0af3c00e0da64ba80d4dbfd0bd76bd129da895db2ed83048d6f535a

Observation ef2d74b1-b261-492d-934c-d8f159487fab · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.464159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:be423aa9026b52cc471484807cfa8caae0cb0c0f607b34d854f7bf60bf9ac90a

Observation 1b11bd3d-be73-40cd-916d-4bf807868b56 · inbound

AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors cites this paper.

AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 129

Resolution
unresolved
no resolver link, observed 2026-07-13T22:49:03.259461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:49:03.259461Z digest=sha256:36ca0960af5af31f2c652cb54ad7638d400103a405d6c6b3617eb01cad7f35de

Observation 6b77c536-6fe4-4752-b714-7494b8aa2b33 · inbound

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization cites this paper.

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.059816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:47:37.605996Z digest=sha256:f069a49533927ffc3ac70abfe3449dde5ae6f8ca703b055e157cb17f26f11af0

Observation 869f6de3-8dd9-4338-8b44-166d6aa071da · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:25.219400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T17:30:36.462567Z digest=sha256:6823908a06954e03ccf183cc6064b75db11db9ed034692852eb640ce6754991b

Observation 84eddbc5-c482-4812-a9d6-7c416a6dbf5b · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T16:38:16.290404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:38:16.290404Z digest=sha256:52f978a42445db42e47152916bf8ef5eb22329b6262652a5e29b775554c0fda4

Observation c194592f-4807-4c2f-931d-091a5e279976 · inbound

PianoFlow: Music-Aware Streaming Piano Motion Generation with Bimanual Coordination cites this paper.

PianoFlow: Music-Aware Streaming Piano Motion Generation with Bimanual Coordination AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:00.065751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:54:30.725820Z digest=sha256:3d797195c29a0ab6d3bd04a861b62afc772f1ff522c29ef6bfcc4107aae20b5c

Observation e4cc6ec4-2ecb-41f5-a7de-4b26bff0e8cf · inbound

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars cites this paper.

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 78

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:20.281786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-10T11:47:42.039240Z digest=sha256:d73ec5c708e5e7a0aed035986e4e62bacf326b623ca1c9750b612b57ab1844b4

Observation edf47450-dea3-49f0-ac3b-54d3b7f004f0 · inbound

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation cites this paper.

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:15:10.370873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T11:13:27.689539Z digest=sha256:5d6b68f0c91fc2fce66791991fe702d79ee7578419b5bd903b0f9cb57c97058f

Observation 00aa68bb-6a1b-4202-a546-18bad7f72798 · inbound

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment cites this paper.

PortraitDirector: A Hierarchical Disentanglement Framework for Controllable and Real-time Facial Reenactment AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:31:04.133105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T03:30:18.645845Z digest=sha256:3322d9a737b1129c98a789b62e1d3cf5a8b071665c45f1b122e0c0997b8fc4f8

Observation 9d3dab7b-01be-4261-b5db-b5f24d0e5bc0 · inbound

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation cites this paper.

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:28.066848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T02:55:09.008954Z digest=sha256:195d2df91e9d2ce375f8e9c3e91ede06897ef8633a75e3df24ab3460eb804d05

Observation 9f3a0d5e-ad8e-414e-bf73-ac6bf6b8856e · inbound

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection cites this paper.

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:06:04.288003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T14:04:52.065878Z digest=sha256:7be801aac82a5bd89ad08990cd83fd1b6df193a998e5bffb16dba5ccb9b1e301

Observation b429ade0-fbb9-496e-8c84-a4209b84db4a · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:11.234958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T20:19:39.565157Z digest=sha256:3a1fcd0cc3a0e8303704ee9b3df1b7a29a3ab7494b0eb71952158f026eb0d666

Observation 8671e9e2-b04c-4b2b-89e0-90fa9bde142e · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:57.546592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T02:18:58.996355Z digest=sha256:ef62822d9806f40d125d83e03a1bf5fcce45c0c13cfd5caa9c31be14deccba88

Observation 3c1edea4-37c8-4276-bb89-6a154abf2069 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:24.949010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:e45cb85e27b5bc283f5cdac386a2aae4cbeeca2bb3df928fecafc62dfd6be183

Observation ebfd67df-04cf-4a8c-b3bc-87646700981d · inbound

Loki: Representation over Architecture for Diffusion-Based Portrait Animation cites this paper.

Loki: Representation over Architecture for Diffusion-Based Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:24:49.856815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T15:23:48.899214Z digest=sha256:87672406329102c6e8673ecb67890b29e2982effb76ce079c9fad48cbbbe9951

Observation 4d174dec-3c35-407f-95c6-8f7806df1f53 · inbound

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation cites this paper.

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.748177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T22:36:53.138354Z digest=sha256:42ab6fdc5883dc976cf9f7e1404c196573ca328f094e5dacd0a3e29e1a370983

Observation 190bd956-2db7-4ee8-a375-ab9eeb6c1bfa · inbound

CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning cites this paper.

CogPortrait: Fine-Grained Eye-Region Control in Portrait Animation via Hierarchical Agent Planning AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:13:27.648129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T13:03:27.312548Z digest=sha256:75b070c8d580995b2a9fbc643fef2c6983c5fd51a5d80cec3855054d41ea6855

Observation 35c1c419-e08c-47c2-a075-f516d9e9a118 · inbound

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation cites this paper.

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.369443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T07:50:56.947671Z digest=sha256:ee3c7e1f71c82622dfbdfe638440af4f0104a48d43116f12f798b7f3b8a353b1

Observation 8294845f-7655-4acb-a80a-8ef13ad32a3b · inbound

Archon: A Unified Multimodal Model for Holistic Digital Human Generation cites this paper.

Archon: A Unified Multimodal Model for Holistic Digital Human Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.783826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T08:03:04.294439Z digest=sha256:49545499c9eb8650c84642dade94bda5eb73963779069e5c4399cf581031cb50

Observation 33df6ba3-c176-445b-8002-068b531b6d8c · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:46:15.580499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:fdd0ed815daede86fd7d15f2f0cdea0ced2db70571fd5421831f2bd3f1669940

Observation 679e79a2-4c15-4b52-8c6a-2414065c442f · inbound

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs cites this paper.

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:06:17.010720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T15:41:59.683979Z digest=sha256:91a259dae2b3cff49a342f2ae40d1805770e199c1afd5b95e51d9160e6998bd4

Observation eba3d6fb-668c-474d-a03d-982536240543 · inbound

Mamba-Enhanced Implicit Motion Learning for Audio-Driven Portrait Animation cites this paper.

Mamba-Enhanced Implicit Motion Learning for Audio-Driven Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.577129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T11:07:20.577007Z digest=sha256:8dec42c1353e497a364b17258be97b4c68f8c7c9c5f67a9b8f4b0786fb55d33a

Observation f3d16bdd-0726-4ba9-91db-abde66eb85a1 · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:40.320788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:590b4af4e057371c3692db71b168cf350a13cd4af214894f7ddd20f77abf1ab3

Observation 2656af68-61a0-4795-8024-6653e9887fee · inbound

ViDS: Video Diffusion Shader using 3D Face Tracking cites this paper.

ViDS: Video Diffusion Shader using 3D Face Tracking AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-31T23:01:31.435147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:01:31.435147Z digest=sha256:9a695ef8ca26cf0a2d5e1f6b6cc08165b6b331b82eda11d1374377856c807845

Observation 4f7baa9b-77d8-4a3e-8cf2-10a4f0f8b209 · inbound

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation cites this paper.

LeapTalk: Breaking the Latency-Quality Trade-off in Talking Head Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T01:32:12.829607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T01:32:12.829607Z digest=sha256:7817e3c8e38fa5803df28f7a8156dce3b13b2b45fb6ab010d3d5e70248e52e0a

Observation 5056342d-4155-41e5-8f60-43ca33456263 · inbound

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation cites this paper.

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T01:03:12.618881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:03:12.618881Z digest=sha256:05bcfd771e473bf0b0eabfe8efad902b457a5832357ed71016f61d6f859eb141

Observation b53fd63d-e3be-45d4-9073-51f037dccf40 · inbound

Foundation Models are Implicit Deepfake Detectors cites this paper.

Foundation Models are Implicit Deepfake Detectors AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T17:41:09.637670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:41:09.637670Z digest=sha256:ee2daafa26b43338fc50a44d428860b24f228904fae234c6c0e5003cf40823e9

Observation 1b2d76ed-79bb-4259-85b8-49a434c4dc5c · inbound

Avatar-Forever: Decoupled Parallel Training for High-Quality Real-Time Infinite Avatars cites this paper.

Avatar-Forever: Decoupled Parallel Training for High-Quality Real-Time Infinite Avatars AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T00:21:12.231872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:21:12.231872Z digest=sha256:3ceb88b2aee2196e674ae1304fb95732c90932480c2e9eaf4026c678c49d01be