Pith. sign in

Paper Citation Record · LEDGER

InfinityHuman: Towards Long-Term Audio-Driven Human

As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 4 inbound Pith citation observations for arXiv:2508.20210.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20210 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:19:37.295188Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:13:11.504526Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:45:40.337791Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc0a63cf-7900-428f-a39b-d6c8e77fb643 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

InfinityHuman: Towards Long-Term Audio-Driven Human , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.050321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.050321Z digest=sha256:59d5f1223ed2db9c4ae8fe29cbcbfb9bcd15bd51195d9ec5ba8ac360ee45897b

Observation 6505ceef-5c88-4522-8425-d8ed5a63c410 · outbound

This paper cites write newline.

InfinityHuman: Towards Long-Term Audio-Driven Human write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.056293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.056293Z digest=sha256:06786fb3809978802e9842011951882b194c58775d3ddc5a8f769f4ba835db75

Observation 271442c1-62b2-4a68-aea6-e7ae265cbd28 · outbound

This paper cites MAGI-1: Autoregressive Video Generation at Scale.

InfinityHuman: Towards Long-Term Audio-Driven Human MAGI-1: Autoregressive Video Generation at Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.062536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.062536Z digest=sha256:e543055e29893100540de14b2f9cc235f19aaf65e658b10318f1a52b1b606882

Observation 36bc65c8-c9c2-4d2d-af15-bb7aad95c0f7 · outbound

This paper cites Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models.

InfinityHuman: Towards Long-Term Audio-Driven Human Vidu: a Highly Consistent, Dynamic and Skilled Text-to-Video Generator with Diffusion Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.068815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.068815Z digest=sha256:d74b99f57fa698546e7c8a52485a48a0955e880acd9c3d1d7c01670740efd78d

Observation 0ed0deee-0350-4de7-94b2-808957e9e457 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.248653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.073887Z digest=sha256:277e60000ee1fd68bf54cb812f53feacc9e756358dc7b25faa42e9d394135d2e

Observation 23e80ff6-605b-4c51-b118-967cc6ce4aab · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.231867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.078799Z digest=sha256:14681e9060fbb3e18c5b724b9f417460c8ae9c4bd1c567afdc4c85f1f12a1f4f

Observation 6ba1d3e8-553d-4672-8be4-43b9c8007171 · outbound

This paper cites E.; Fang, Y.; Lee, H.-Y.; Ren, J.; Yang, M.-H.; et al.

InfinityHuman: Towards Long-Term Audio-Driven Human E.; Fang, Y.; Lee, H.-Y.; Ren, J.; Yang, M.-H.; et al

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.215824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.083837Z digest=sha256:7e150ac704acc7601670b25a64e411695327e07a6df0104661e4edc782412c5d

Observation 69e9fe5b-c627-4661-b3b4-3e72b6611d20 · outbound

This paper cites HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters.

InfinityHuman: Towards Long-Term Audio-Driven Human HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.088684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.088684Z digest=sha256:30632a385b01b4d912c663b69119018d7074ac6b2cf91ff66aaf107d7d2f8340

Observation af6ce6b3-4224-4969-821f-bcbbb626aca4 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.199900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.094997Z digest=sha256:54dcd8b15a024017c6651dfa6c19db800aaddc4c2f271086fe43ae66adf3fc27

Observation 038f89b7-c932-4c4b-8e8e-2c3d2d0df5ea · outbound

This paper cites S.; and Zisserman, A.

InfinityHuman: Towards Long-Term Audio-Driven Human S.; and Zisserman, A

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.183644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.100643Z digest=sha256:7e2564a1535205dbe5dbb8a248fa9c9f24ceb7fd38e787d22335bbc163dbf675

Observation c85cec5c-d3de-468e-b6cd-8fe99dcf3647 · outbound

This paper cites Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer.

InfinityHuman: Towards Long-Term Audio-Driven Human Hallo3: Highly Dynamic and Realistic Portrait Image Animation with Video Diffusion Transformer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.110895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.110895Z digest=sha256:9122a90ec384d080430b4133100565bb610d1693fcbc74e5be8fb886b720a912

Observation 2a2ed21f-c783-4b36-8408-d28c9f114b26 · outbound

This paper cites HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.115944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.115944Z digest=sha256:be9fa26c0e719b668209eb3c037837a6f002ff5af637bae9c53c03a7e79b0868

Observation fde95110-063c-4cbc-b40e-69b6b689c989 · outbound

This paper cites OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.120737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.120737Z digest=sha256:7075d4dcfc3234ecb07ba9e4d9a1b5689a72d162e1daff7471221c9752d50512

Observation be44d807-5be1-4d44-ab87-3bdd660a9c7b · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

InfinityHuman: Towards Long-Term Audio-Driven Human StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.125808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.125808Z digest=sha256:26629667d7e0d9434f77c4edef85489fac2f487759969fdd9f878921ec9e9e9e

Observation 5d6bbbce-33c5-4d1f-a207-93c7d5aa1711 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.130452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.130452Z digest=sha256:fe5821af3bd553b550d441fa16a3326502218e15e26a14278a3e66d4d99a5bcb

Observation 74d7fde3-0a2b-426a-b704-d1d1e071aba8 · outbound

This paper cites Classifier-Free Diffusion Guidance.

InfinityHuman: Towards Long-Term Audio-Driven Human Classifier-Free Diffusion Guidance

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.134859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.134859Z digest=sha256:58cd2fcfbd79b8932487925eeef64fc6d3a6c01eb57488cbb85d474b81e90301

Observation 2f00f6c6-edfe-4e6b-a774-de02a241b6e2 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.156136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.139672Z digest=sha256:c5bc3e951437bbc36adcc08701c2dd9813dcb50d62954319e4c15b4aaaf6341d

Observation 16202775-438d-4101-9eef-a2e57bc02574 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.144059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.144059Z digest=sha256:190e237439f34e3fbb2585d3143be12a75a225da9c812e395537de54916e61d4

Observation 8ef05fc4-daad-43b3-84d3-de7da141afac · outbound

This paper cites ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving.

InfinityHuman: Towards Long-Term Audio-Driven Human ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.149071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.149071Z digest=sha256:e30d1be5bcfbe7353d81c69f32a36c0c4cc20be5fb696ec422300197281338a5

Observation 146a4c50-26b6-4f83-a965-e321afbfb3c1 · outbound

This paper cites Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency.

InfinityHuman: Towards Long-Term Audio-Driven Human Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.153799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.153799Z digest=sha256:372101d137cc778b928a1b6165bafdfef71b92d1389a22d6f3a3cbd8963a48a8

Observation 469de640-625a-4dff-b7be-60a476dd0985 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.140059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.158574Z digest=sha256:8611b4d5c4badeb24913184bb636ff9891b5ba39c06c38961c239408d784b345

Observation aefc38d3-8410-471e-ae5e-979f2c89fec7 · outbound

This paper cites Sapiens: Foundation for Human Vision Models.

InfinityHuman: Towards Long-Term Audio-Driven Human Sapiens: Foundation for Human Vision Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.163455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.163455Z digest=sha256:55d1d6c045bf435bffa20f88405cbb027b363d20a623911ab7510ba4781002e6

Observation 2b649be8-018a-4de3-b28b-05f9b18454b2 · outbound

This paper cites Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.168237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.168237Z digest=sha256:204b16dc681b264a38dfeb2b85104e0aa35b8b78ffbf6432ab8124c64dc496dc

Observation 3022cb00-cc3f-403b-b724-35d80d10d2ff · outbound

This paper cites CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention.

InfinityHuman: Towards Long-Term Audio-Driven Human CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.173088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.173088Z digest=sha256:122c9c43c9a121bfb5112a5b4c8b35d0e2af469d73a0f5a25aa91c41cd974726

Observation efb1e719-f16c-46fd-9d6d-61eceaba2a69 · outbound

This paper cites OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models.

InfinityHuman: Towards Long-Term Audio-Driven Human OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.178523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.178523Z digest=sha256:fe6ae4ea50b52e0ab98903bc1bbdd7de8ce419f4078399c00114765ddc93e819

Observation 20f435bd-b508-4dfa-addd-d92d80bb6b1a · outbound

This paper cites Flow Matching for Generative Modeling.

InfinityHuman: Towards Long-Term Audio-Driven Human Flow Matching for Generative Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.183577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.183577Z digest=sha256:5a43e9731b43d9789d0c3abcbae9264276356e223883c731de93da0ce86a2476

Observation 202e6d05-139c-4a7a-b6a7-0f719cdf91d8 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.188569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.188569Z digest=sha256:113317595d0fe536c37acc9e442595a4a044b67e54449450064ed2ca0f8649b1

Observation 2718cb59-cc64-459f-b56e-103e47549e3a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.125171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.193017Z digest=sha256:eff6ee06b8317a35a7f9fc861808894f8bbad1079419cb87c1822de6131c9a10

Observation d7d657ff-ebc3-4af5-8c10-adf6cf5f7986 · outbound

This paper cites FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling.

InfinityHuman: Towards Long-Term Audio-Driven Human FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.197844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.197844Z digest=sha256:bac91e52d732b0ae972a45a58339085a6c8fba7f8ec8582758314eabe51c11b3

Observation 97676dc0-2473-447e-b5e5-e0efa5e7aeac · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.109624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.203126Z digest=sha256:da18a3947493bdec6db393cde7d7b9f48f4ff16543e08022d54b392e28750023

Observation c1f3cd11-6ade-422c-b551-322c9f20e18a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.093372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.209356Z digest=sha256:98a4451173407d5be0fb5e7d9f1681b0a8d72e7c238abf0f6f8a49264f1a66a9

Observation d34993e8-472d-4343-8392-b985e2748253 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

InfinityHuman: Towards Long-Term Audio-Driven Human Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.213971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.213971Z digest=sha256:d691064ae5bad9e42de8922517c4fc36e86fe88fff7705965690f9ae04bce407

Observation 272fa9aa-e990-4612-add2-f17b4de8a0d7 · outbound

This paper cites V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation.

InfinityHuman: Towards Long-Term Audio-Driven Human V-Express: Conditional Dropout for Progressive Training of Portrait Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.223978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.223978Z digest=sha256:da16fe9eb817150900279135a954a6e99a61a8b5e379f4c464310e3c4aeb2f96

Observation d25195cc-2719-4df0-8b08-247ecd523f08 · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

InfinityHuman: Towards Long-Term Audio-Driven Human Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.229734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.229734Z digest=sha256:fb3e82a775426b479f40359399c7b7b1d8d61259a2647adee807c5c468b9a257

Observation 62331e57-43f5-4b1b-a284-cc7be6a613b5 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.077176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.234878Z digest=sha256:16642558a89c94604264751a068d406c35f0388d63e23a827bdd005d86ec9993

Observation 5cf5dd96-bd4c-4fee-82a1-fe661de31133 · outbound

This paper cites FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis.

InfinityHuman: Towards Long-Term Audio-Driven Human FantasyTalking: Realistic Talking Portrait Generation via Coherent Motion Synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.239985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.239985Z digest=sha256:80e9c2a01cba883c08af4576cf7495d677d7d095a125d5f314c0c41e616259e1

Observation 2728a245-1ee2-4b8b-9c40-ab9ca16ade0a · outbound

This paper cites AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.244894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.244894Z digest=sha256:77730e5e36a6753eed2607202a98daba34b061dc435166cb7f0d2b682cb2ac60

Observation f037a761-4b70-49a6-82b6-336aaf4a3a45 · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

InfinityHuman: Towards Long-Term Audio-Driven Human Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.250325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.250325Z digest=sha256:6b02f8e3aa64d2b50c28d4a8a4606abb697951dabb079303f38f6671786dd68f

Observation 38bb490f-c1fa-4a8d-83aa-6c91a5e74898 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.255149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.255149Z digest=sha256:3e028620861a0d675f1f17f5d6641ed4d3334b39dfea91bcd3ac218fd05c60da

Observation 62b552ec-c597-4ea0-8810-5b3f829d2845 · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

InfinityHuman: Towards Long-Term Audio-Driven Human Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T15:19:37.259854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:19:37.259854Z digest=sha256:c24db42d58d73e5060b97bb060aae7a5aaef8760f2e35267928c8f13180458d3

Observation bc0897b5-7993-4954-b38f-c6b837c17fcf · outbound

This paper cites T.; Durand, F.; Shechtman, E.; and Huang, X.

InfinityHuman: Towards Long-Term Audio-Driven Human T.; Durand, F.; Shechtman, E.; and Huang, X

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:19:38.061494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.265364Z digest=sha256:bea399972c4f1d95faa65c35f0d18a60793c71bf83982a52d54acfe37530aa73

Observation aa22d217-42a1-4175-86d8-d5c0af9172a9 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.045258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.270656Z digest=sha256:ab6850800b5eb4f0d1f730f85839013d9ee08e5303fa96b082fc79aa4adb9b1f

Observation 712e5916-70f3-47b7-ab08-bff605e41d50 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.028130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.274971Z digest=sha256:d69344695665ae4b0afcb50fe35934cf614b0190acdfb3edf37ea047c4281a21

Observation 5457650e-b5c7-4784-a037-f3279b68171a · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:38.013438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.280329Z digest=sha256:56b22e1a51647f2cb33ee5ffade5744799dfe11e673711a26783f1597f7c825c

Observation 72e7c357-8f3a-4075-abd4-21cdf2c8c831 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.995922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.285521Z digest=sha256:23d330fc255a1fc6e2c1227435c1b84488170e9463e43e6dafd679bd0db0b114

Observation 5a86f957-c397-499e-a732-7b44102fb00f · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.979057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.289764Z digest=sha256:fe906e20c93f57eec2afeb698989c1f4fcac9a4d5514b2e9c5dbc05ab2236b3c

Observation 7ff90854-b684-44a0-9008-1925e1374859 · outbound

This paper cites an unresolved cited work.

InfinityHuman: Towards Long-Term Audio-Driven Human Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:19:37.961578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-05T15:19:37.295188Z digest=sha256:11029a3b628599ae118989e3f43e874a356bdb1ca0ccd1ef1a5e45251e5e173d

Pith citing papers

Observation 8cc9086a-b1ed-47f3-9a91-3f6e7527df13 · inbound

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body cites this paper.

ViBES: A Conversational Agent with Behaviorally-Intelligent 3D Virtual Body InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:08:36.256770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T22:04:07.403410Z digest=sha256:b4c0a9ec625dfef90b3e1a6d7b0dcce13f3f5261a35a1c0faec18b3d3a4bc2f6

Observation b2c3cf4f-83f6-44c6-af86-5a5b00ca820b · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:20:22.355150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:eaa0010c586b2d329f764fb4f3de2a0a75e647b574ae198065d0d17929231ee7

Observation ea9b4d82-3dad-4742-a484-16aae52bc771 · inbound

Generate Your Talking Avatar from Video Reference cites this paper.

Generate Your Talking Avatar from Video Reference InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:29.794020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T05:32:04.519820Z digest=sha256:f6aeb490341f53298423d42352c97eec061f0914116a0676d1252d25dd5eeb14

Observation 2200ada8-547f-435e-8314-93b52070d8c4 · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation InfinityHuman: Towards Long-Term Audio-Driven Human

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:40.339046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:0ea0f89c0e3a2d1fcb4a9149ed61c819857191a2458174738f2d33c081399626