Pith. sign in

Paper Citation Record · LEDGER

Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2410.07718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.07718 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:26:50.715253Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T21:46:15.594455Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation af259182-f475-49fb-915f-2c93f8bd620b · inbound

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation cites this paper.

JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:03:12.594476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T16:59:57.727035Z digest=sha256:38318080440a30e64ec013593e5609cc78d67e46c406edce364e77417b0ef38a

Observation 540aa358-8131-4303-a544-84b8682291ea · inbound

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation cites this paper.

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:23:15.413183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T17:19:51.411937Z digest=sha256:fb4e680d366368648a30ba83dbbe968beff8763702465776e467be1d8b605414

Observation 9609bc2f-afc8-4286-a2df-7655518844d9 · inbound

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling cites this paper.

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:42:43.835012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:42:25.803856Z digest=sha256:73640286ad68e04e805701909cb72b10d69ffc0a7f99703429e37f166c827834

Observation 0b127765-b937-43a6-9ef4-c84c1a3a5510 · inbound

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits cites this paper.

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T16:30:32.718987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:30:32.718987Z digest=sha256:1ff651c051a2f66ea0c6d210a3c2c83de962408f6fec948bca8f8e9de4f12797

Observation ae42507f-bdfc-4b0b-bbe1-d07c187da44e · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:26.836265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:30:36.462567Z digest=sha256:4d43dc85514f61d2ab2411a16f916d8e2a7fdaedec0829ac20cf24ab31d2c80c

Observation a82d69aa-9a65-4f55-8e25-a7a934cdff73 · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T16:38:16.042209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:38:16.042209Z digest=sha256:a394d7cd6a4d702f53c8a4466a2db6dc3a04da3b10dcc479e9678d46cdf509c7

Observation 811af94f-3d20-4d2a-83f4-1265b283c766 · inbound

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction cites this paper.

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:28.578744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T22:06:58.480052Z digest=sha256:53d9a148bf40664000536e42746e424ac910710a97f6aa2a81b0cab3428c1aa1

Observation c1750b6d-025e-4cae-9526-91bc8104f0bd · inbound

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence cites this paper.

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:11.198187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T08:27:41.839123Z digest=sha256:b0995fe1f9d126a074d3e557329417531e873458fa83dfb3003e063e90f9df32

Observation 6e8726f5-ca6b-4469-bbdd-86eb19d08c1c · inbound

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling cites this paper.

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:14.180854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:44:36.000353Z digest=sha256:9ad56b755f8a9a6b1f79b5c77108cbc4be43ca98b92a664bbe1d501c04124439

Observation b82dc350-6589-4fd9-ae47-fee333ec6da1 · inbound

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation cites this paper.

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.812027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:56:19.795651Z digest=sha256:78c7d967a1a549663aa696b198ea86624fbe86786f3083943a9c9bf4b354437f

Observation 2fe91509-114e-4f7b-b718-75bea44c002c · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:11.446098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T20:19:39.565157Z digest=sha256:6623badaeaf9d5ad8656a8a41174259e8391429311a4c8e6c74181cfc567542d

Observation df2841fd-a245-42a7-889f-8565b1735a39 · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:57.382344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:18:58.996355Z digest=sha256:20f23ba803042ffed422c42ffa5c966b7e1ab9c38422a17e1fac9d99d6075dc4

Observation ae3e1715-0ae4-4eb5-ae40-1be6464048d9 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:24.973274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:825196c1a7e4b1c1c3f76bd18f2784c8689e7286169ea017bb0c40197a06050a

Observation 04324c47-4557-4776-a937-9738122cda53 · inbound

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation cites this paper.

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.682150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:36:53.138354Z digest=sha256:7f7b7c7ef703c1f2c2a9845707a2bf08a960ac61d68d57d41fd8a9910abb4d20

Observation f94e1102-ae94-45f9-a924-002237519f39 · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:46:15.595923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:699594884d73a9286ad6481e4546ee90845b9025426632eb128ad04dbfb8d77a

Observation 5ebf10fb-1614-4105-a973-0a5927c1d4d1 · inbound

OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data cites this paper.

OmniDance: Multimodal Driven Dance Video Generation with Large-scale Internet Data Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:44:19.601000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T06:35:58.943585Z digest=sha256:a233593cc46d6f84fd6b44af0038c5db835dca71957871db44783492c4982bbe

Observation 0f139d01-4afd-452c-a170-dc951abaa70b · inbound

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation cites this paper.

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:35:40.322410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T06:26:20.283349Z digest=sha256:62d74e0cf9a314bfad023e0a09f8dc61c7183d4cf502dba311346761b8795bb4

Observation e92832fa-9631-44af-a48d-dd43c8172355 · inbound

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation cites this paper.

Towards Flexible, Natural, Efficient Interaction for Conversational Talking Face Generation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T09:27:55.261808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:27:55.261808Z digest=sha256:53a3e31ac037e05c88e88229672530230627bcc36a67411953a24fc681a83927

Observation 73216a48-287c-4f7c-90d1-e3f5260b229a · inbound

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head cites this paper.

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T09:02:19.214740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:02:19.214740Z digest=sha256:6b30af60df42c95403648f9994bcae7118a9c1857bb2542f8d9decc8fdc005f6

Observation 3b449691-410d-4e27-a27a-9c37f1d6f9d1 · inbound

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation cites this paper.

Proxy Avatar Meets Low-Rank Caching: Real-Time One-Shot Emotion-Controllable Portrait Animation Hallo2: Long-Duration and High-Resolution Audio-Driven Portrait Image Animation

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-04T17:26:50.715253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:26:50.715253Z digest=sha256:e12c66529d2813b0f7dd9220622c77d74800be3bc3381d8728aa07e77e5863e6