Pith. sign in

Paper Citation Record · LEDGER

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2406.08801.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08801 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:36:52.656632Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:06:16.996788Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 454d8982-09dc-40b5-b107-adef64a65c38 · inbound

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation cites this paper.

Multimodal Diffusion Transformer with Memory Bank for Scalable Long-Duration Talking Video Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:23:15.339474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T17:19:51.411937Z digest=sha256:92a8553901524f7eff7ef7ee41ca12284c8cb1242293fade6dfff013869cf6b8

Observation c1f186e3-bbdf-4141-bd1b-e49110f87c7b · inbound

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching cites this paper.

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:50:51.078255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T00:46:39.196042Z digest=sha256:1285223faaea757ad0c642e40d9970aa656320e72a86a7466a83dd0ae4a88f38

Observation e0f06b7c-dffa-4f88-a39b-534522cb11ef · inbound

Human Motion Video Generation: A Survey cites this paper.

Human Motion Video Generation: A Survey Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.656632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.656632Z digest=sha256:a2cc99c1bf3c6c69ce688aca519fd16b2f7448b6df4ba7934f21e2a852e8c823

Observation d0d0c22e-34d9-4ad0-9c96-9e2eb28e041c · inbound

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling cites this paper.

FluentAvatar: Flicker-Free Talking-Head Animation via Phoneme-Guided Autoregressive Modeling Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:42:43.789808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:42:25.803856Z digest=sha256:2e4bb29399ff30d89ea6945ff99cb3413a783d5e460e107c81861eaca6df69a0

Observation 697de81c-4365-4064-9772-1108bc316543 · inbound

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits cites this paper.

STARCaster: Spatio-Temporal AutoRegressive Video Diffusion for Identity- and View-Aware Talking Portraits Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T16:30:39.500409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:30:39.500409Z digest=sha256:d029828917a0397a27c13e643f164f695e94c80b9a1e4e32ff3aa22d8400062f

Observation 17b76d59-bf52-4010-a9e8-b08a79c75706 · inbound

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization cites this paper.

AvatarPointillist: AutoRegressive 4D Gaussian Avatarization Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.152270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:47:37.605996Z digest=sha256:1ccf1012e7dbf089695e39481cd97d70f9007a83cc3a2c7e6a9ad75a72b535f8

Observation 0e727eef-15e8-4aac-b21d-edae96a3c738 · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:41:25.776975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:30:36.462567Z digest=sha256:945c80e02018436f07eaf0ed814eefb739aa4de45ea0a5a2872b178a6264b26b

Observation 205693a5-d40b-47e9-9cea-eadd417d5a37 · inbound

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation cites this paper.

SyncBreaker:Stage-Aware Multimodal Adversarial Attacks on Audio-Driven Talking Head Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T16:38:16.303378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:38:16.303378Z digest=sha256:35acf8395fc5f94023d229b2650df8e6c761971138cca3ec2dd2192072423900

Observation b7ead486-41cd-413e-8ce4-48e880f4463d · inbound

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels cites this paper.

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:01.763121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:16:53.603971Z digest=sha256:8655c118c7fe786fe5474af9f588807c6a5d3713ce5d694179eb284f5c0d206d

Observation 898331d4-11ad-4304-b41f-de82d2d0d867 · inbound

Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection cites this paper.

Direct Discrepancy Replay: Distribution-Discrepancy Condensation and Manifold-Consistent Replay for Continual Face Forgery Detection Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:31:01.252071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T14:50:39.077649Z digest=sha256:c365a18ebf4010f50210632df45e8ef3df20a21b26f6874535131f4878abdf2e

Observation 1b4cb1a1-22f5-4bab-8424-996a9907d853 · inbound

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars cites this paper.

Giving Faces Their Feelings Back: Explicit Emotion Control for Feedforward Single-Image 3D Head Avatars Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 84

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:20.275679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T11:47:42.039240Z digest=sha256:9fe921f57891cb39cdddaedb7f195ac30ca17aaab99fd1b3829ba9b9f899a359

Observation ed2442b5-6753-4c39-b7af-53b558c18d0d · inbound

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation cites this paper.

TurboTalk: Progressive Distillation for One-Step Audio-Driven Talking Avatar Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:15:10.394608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T11:13:27.689539Z digest=sha256:7ea0a475556811ae544b417f7e29c467c8caa501ffc94cdffc8025228fcf3f49

Observation c6fe6c3d-71db-4fad-9345-e425bcfd5ca5 · inbound

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation cites this paper.

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:28.097078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T02:55:09.008954Z digest=sha256:a976e6557db3cfa8183eb9fef46573a73bde07842daf5c63caa1ea1d5723968c

Observation c2658e7b-c65f-431d-be8b-547f74962c57 · inbound

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction cites this paper.

MeshLAM: Feed-Forward One-Shot Animatable Textured Mesh Avatar Reconstruction Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:27.786083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T22:06:58.480052Z digest=sha256:0d7802d2784bab34cd967505e6705d8be1ae9c4bfa55554f4daee0c69999dcdc

Observation 6178aacc-2956-4397-afa6-f704ee068401 · inbound

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling cites this paper.

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:14.299885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:44:36.000353Z digest=sha256:8f949fb59aa3d8432351e675ecafafd49c520f2bd54fcfb7dc984998950b1f8b

Observation f0d3deec-29e2-4350-b3ab-4b1dd035a100 · inbound

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation cites this paper.

Hallo-Live: Real-Time Streaming Joint Audio-Video Avatar Generation with Asynchronous Dual-Stream and Human-Centric Preference Distillation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.720703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:56:19.795651Z digest=sha256:7967d4ca36e3e392f9b4620645f8adfc5819773838c2d70f08f8c4b8b739a71f

Observation a13e0a97-4797-4d0e-95fa-2842b1557e0b · inbound

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations? cites this paper.

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations? Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:14.711337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:42:31.456648Z digest=sha256:9037923a845514bafa0a578ae2192a8e2fcc88ffa2f4b928f8d98d874b22ca79

Observation 91275adb-8fef-4224-9e85-f1acd0d5c7ad · inbound

Generate Your Talking Avatar from Video Reference cites this paper.

Generate Your Talking Avatar from Video Reference Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:29.833106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T05:32:04.519820Z digest=sha256:4095c4769061c23d8807f2667c0452138cb2b6232e4d226d8f819bf3b7773bd4

Observation bd95ef94-b0d9-4823-b202-c18502e0671e · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:11.240985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T20:19:39.565157Z digest=sha256:4b4cebbe8cc35fbb74a056635307919021b335eb485d66c6e9ca43a631464282

Observation 93601e6a-008c-4ed1-b8af-6296ce99a58b · inbound

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation cites this paper.

AsymTalker: Identity-Consistent Long-Term Talking Head Generation via Asymmetric Distillation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:57.523880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:18:58.996355Z digest=sha256:d8c288208a34dbaec1647a2bce4ad349b0e6dc2aea748fc8372f0cecffa25f2c

Observation 09b818d7-6573-4f55-936e-f1ba2c36e39e · inbound

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency cites this paper.

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:51:08.915687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T13:30:47.547011Z digest=sha256:f4a2325723b35c514732fe0dddada337063b1769eb8679375566cef8890b9f6d

Observation 86e8d2fc-022b-4df2-8423-3a3f04da3434 · inbound

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency cites this paper.

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:54.099448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T02:10:23.616363Z digest=sha256:b6b7fe77b9aa90e64ebb0d1becda7d748de3d5d94cfa761ce41ac83c10b09e26

Observation cae1bfc1-313a-4c5a-abd6-e0c6a72ec99c · inbound

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency cites this paper.

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:02:27.640716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T06:58:57.520289Z digest=sha256:d6f248fa99dcfac9fcc94b66633ed7ddb27a890d546ff91147487b64305ba6df

Observation 92272046-ae8d-4b36-a705-39063cb981dd · inbound

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency cites this paper.

Eulerian Motion Guidance: Robust Image Animation via Bidirectional Geometric Consistency Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:25:44.775363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T23:25:13.965611Z digest=sha256:6efafcefd43df39f8285df816ca6a177d74637e3a8373bdf5cf072151a8bfbf6

Observation 10a4157b-6876-4985-9995-0de49e6314b6 · inbound

HighSync: High-Quality Lip Synchronization via Latent Diffusion Models cites this paper.

HighSync: High-Quality Lip Synchronization via Latent Diffusion Models Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:17:48.650577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T21:14:25.606097Z digest=sha256:950fb9fdf277f6fc01fc69092173041ab392b5f2c44eaac339b85de21a0356c0

Observation 08535c7e-7e6c-4f05-835b-ab2fd1ed2a52 · inbound

Image-to-Video Diffusion: From Foundations to Open Frontiers cites this paper.

Image-to-Video Diffusion: From Foundations to Open Frontiers Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:08:24.871201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T15:06:02.084336Z digest=sha256:25a777c59853e5d832788b4368d6b20efa6a4ada1f85ada91a7c49fc87c04df3

Observation f34ed0c2-6d49-439c-9a29-b38f52266d0d · inbound

Loki: Representation over Architecture for Diffusion-Based Portrait Animation cites this paper.

Loki: Representation over Architecture for Diffusion-Based Portrait Animation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-30T15:24:49.848910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T15:23:48.899214Z digest=sha256:03593d0907aa8cb586ea05a5c2c02499ed332621093044c64978450bac07b62d

Observation 4d58aa8a-0465-4083-80dd-4555587343bd · inbound

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation cites this paper.

Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.687595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:36:53.138354Z digest=sha256:645da4e33c8d118b35eb448a2ba6273f76ed4955f10c0f773e8504b63e8f729a

Observation a89fcbd1-5187-4cb5-abf3-3dc878572ccf · inbound

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation cites this paper.

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.366896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T07:50:56.947671Z digest=sha256:2058ac366c09a9ddc0a6b5b37a6a6a1a867af2832476af9f1cacce7e8a60351c

Observation 210e1107-9f6b-4bd8-b811-d59d4911bc8e · inbound

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation cites this paper.

Temporally-Aligned Evaluation for Audio-Driven Talking Head Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:46:15.588139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T16:19:55.048831Z digest=sha256:50202f7cb66906537dcf45078dfecc5167f72979e9f6d95b2faf369f98890947

Observation 1a0bf6b9-050b-4aae-9500-3c2f8ad4609f · inbound

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs cites this paper.

Real-Time Generation of Streamable Talking Portrait Video with Reference-Guided Deep Compression VAEs Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:06:16.998665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T15:41:59.683979Z digest=sha256:fc3213fab01a2bb018c09a7d6a13238b4583e8d0cc6b551019f0890be9480873

Observation 9344f1e9-3dae-430c-a33d-06ccf449486f · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:40.302530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:ee17f04c8c7b056f98c29a9650e75c8c07cf394ea19fb2a426dc2a4b9d7a3430

Observation b493bb94-9218-4ea3-8a09-b0dfb81ca557 · inbound

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head cites this paper.

FA-LAM: Focus-Aware Large Avatar Model for One-Shot 4D Animatable Gaussian Head Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T09:02:19.386267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:02:19.386267Z digest=sha256:fe3e8d2d8376616d2c2819a8c8fe623d24784c174390d461b10aff73cb362b5f

Observation 6038c66d-f163-47bc-bfd3-a30d05155090 · inbound

ViDS: Video Diffusion Shader using 3D Face Tracking cites this paper.

ViDS: Video Diffusion Shader using 3D Face Tracking Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-31T23:01:32.039915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:01:32.039915Z digest=sha256:88eec2370ea6f9c89636fdc61ee14fcbefc0a06008dcdfac83e751502fe3dca4

Observation 44928e9b-9bca-4652-8a70-7e1e460e7b88 · inbound

TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation cites this paper.

TaoMate: Anchor-Guided Memory Bridging Evolving and Reference States for Real-Time Audio-Video Digital Human Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-31T17:08:14.198579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T17:08:14.198579Z digest=sha256:6b789e6b93e281314cf285cdc55b7788334e30943e00454e1724b05e66ee90e8

Observation 8192666c-149d-491b-9cfc-40000194f655 · inbound

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation cites this paper.

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T01:03:13.201630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:03:13.201630Z digest=sha256:2549c9f85e9316e69a78e94bd4061c5ac71818a02b2beb42bf1b561ad51dbfcf