Pith. sign in

Paper Citation Record · LEDGER

Human Motion Video Generation: A Survey

As of 7 August 2026, this Paper Citation Record lists 100 of 228 outbound references and 0 inbound Pith citation observations for arXiv:2509.03883.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.03883 v1

Coverage vector

measured 100 of 228 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:36:55.691493Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 228 outbound references displayed

  • verified exact9
  • verified fuzzy0
  • unresolved91
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f95da5f0-5166-4cbc-91b5-4003ab12495c · outbound

This paper cites Deep video portraits,.

Human Motion Video Generation: A Survey Deep video portraits,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:48.887967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:48.887967Z digest=sha256:d3e1141a7dfd3dd75aaa34682d21a93a63d0c7444a417b7834fd3e24f73a8dd3

Observation a24baf39-178c-4360-adda-a78ea16f0c88 · outbound

This paper cites Faceformer:Speech- driven 3d facial animation with transformers,.

Human Motion Video Generation: A Survey Faceformer:Speech- driven 3d facial animation with transformers,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:48.970520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:48.970520Z digest=sha256:b25ccbbe11d4fddae533dd73c69caafac5518dc463adf86c3f4a78cc389d1ca6

Observation cd7f49ad-3ec6-4700-bd27-747d8cefa3a8 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Human Motion Video Generation: A Survey Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.092801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.092801Z digest=sha256:5f6311ca963c73f8d7ae3991d165cf1bda0c2980970920d3f38333cab3aa3dca

Observation 26813512-657e-4181-b184-6d7aa6204fa2 · outbound

This paper cites Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,.

Human Motion Video Generation: A Survey Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.159879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.159879Z digest=sha256:f2f2fa03053d2c7c81f7f61e797d84d255907eec4dfbe13c070e9c49d52197de

Observation 94db31dc-7003-4a00-8710-6af5c5cab0aa · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Human Motion Video Generation: A Survey Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.239210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.239210Z digest=sha256:4ae2ada8f6b5e26549e3a6e861f3b0fa15aa1d8b598869c220661100b969ddbe

Observation 7b169ad2-017c-4d24-8270-8e3568cb2bfa · outbound

This paper cites Faces that speak: Jointly synthesising talking face and speech from text,.

Human Motion Video Generation: A Survey Faces that speak: Jointly synthesising talking face and speech from text,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.318463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.318463Z digest=sha256:1403e30325ac1309e8dc1c79c5c6ef74988abb262460ca63163977dee84d6ba1

Observation 1cef78bd-e8c1-4cf2-9800-3447126b7a23 · outbound

This paper cites MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion.

Human Motion Video Generation: A Survey MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.374447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.374447Z digest=sha256:3c7f77bbe5cc6413e83cac63f234d8b874559149301e348e97bd7ff79374145b

Observation d40138a0-a809-478d-b4d4-c46d4ed8ac14 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis,.

Human Motion Video Generation: A Survey Nerf: Representing scenes as neural radiance fields for view synthesis,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.451505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.451505Z digest=sha256:2a33d628e3b6ecade58753e262054c10c3db5d092e6cbd06bd241c0f4e30dd68

Observation 7b7fd77b-ed8b-4cef-9f87-1a6789fa5f00 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering,.

Human Motion Video Generation: A Survey 3d gaussian splatting for real-time radiance field rendering,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.513818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.513818Z digest=sha256:6e8dc7570758601780c3bd4a755c477a09924f08062687229acc451a2e2b6c0d

Observation 7eeb9952-bb2b-4d3b-b5c7-7f1c41679862 · outbound

This paper cites Deep person generation: A survey from the perspective of face, pose, and cloth synthesis,.

Human Motion Video Generation: A Survey Deep person generation: A survey from the perspective of face, pose, and cloth synthesis,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.591482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.591482Z digest=sha256:4dc5e31cdd63349af90746748897085cd49b58d62d25ddcb8fb7e6f7a478246b

Observation 294a4644-7e29-4b56-846e-aeb8949c0813 · outbound

This paper cites A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights.

Human Motion Video Generation: A Survey A Comprehensive Survey on Human Video Generation: Challenges, Methods, and Insights

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.645540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.645540Z digest=sha256:a9ef487be35fa385e3cea53996cb3fa325b58d308d3209d809e64c8ad9cbabe0

Observation b01d6abe-af0b-4308-b541-213e8c2377e1 · outbound

This paper cites Image-based virtual try-on: A survey,.

Human Motion Video Generation: A Survey Image-based virtual try-on: A survey,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.742383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.742383Z digest=sha256:4b6bfe51e68008fdcdba80b164ceadd81b3c710b6cbf65edba14da73b2438484

Observation 52958468-5877-4275-ae6e-7b9dcd5dec0d · outbound

This paper cites A Comprehensive Taxonomy and Analysis of Talking Head Synthesis: Techniques for Portrait Generation, Driving Mechanisms, and Editing.

Human Motion Video Generation: A Survey A Comprehensive Taxonomy and Analysis of Talking Head Synthesis: Techniques for Portrait Generation, Driving Mechanisms, and Editing

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.954166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:49.794534Z digest=sha256:a3da5993fb13dea4ce36fc10f099695f40a747cfcde9815e5199fdad71fe431f

Observation f829e0b9-2e0e-4499-bada-7ba5bb7aadbc · outbound

This paper cites Multilingual video dubbing—a technology review and current challenges,.

Human Motion Video Generation: A Survey Multilingual video dubbing—a technology review and current challenges,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.872986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.872986Z digest=sha256:a5870e20a5598143cb315a79c4e3dc441917458ab8b0b575c1f6f086d9e76788

Observation b3705933-00bd-4c0b-8dea-af99e0cac63f · outbound

This paper cites Difftalk: Crafting diffusion models for generalized audio-driven portraits anima- tion,.

Human Motion Video Generation: A Survey Difftalk: Crafting diffusion models for generalized audio-driven portraits anima- tion,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:49.948773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:49.948773Z digest=sha256:4743f9cfde60c4f423124e572c466a294eaabc8b7fe43b64fefb73079bdf8ce0

Observation 20c57b9c-547e-4543-ab8f-cb9dbbbb4769 · outbound

This paper cites Identity-preserving talking face generation with landmark and appear- ance priors,.

Human Motion Video Generation: A Survey Identity-preserving talking face generation with landmark and appear- ance priors,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.025023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.025023Z digest=sha256:429a8ce48bd7c0d3536fcd08abe0dc5ab5821613aa70061c238aa2411c88f285

Observation 1b2b5664-056c-4e89-9ed6-3158cdefa1af · outbound

This paper cites Affective Faces for Goal-Driven Dyadic Communication.

Human Motion Video Generation: A Survey Affective Faces for Goal-Driven Dyadic Communication

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.075167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.075167Z digest=sha256:2f092e79ddca8d252e7f0ad2b96ed65e7eac257aa7f0aeff8a66d379cb6d5633

Observation 772294c5-ddd2-4528-8f9f-a8cd4e122bd3 · outbound

This paper cites AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents.

Human Motion Video Generation: A Survey AgentAvatar: Disentangling Planning, Driving and Rendering for Photorealistic Avatar Agents

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.913967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:50.144987Z digest=sha256:e1a0ed943f56d75323104d25e997a2fb4eb1f5636d657890e2ed8daf0e25cd60

Observation 9a49104e-7b28-46f4-b392-fc81ff03714d · outbound

This paper cites Instructavatar: Text-guided emotion and motion control for avatar generation,.

Human Motion Video Generation: A Survey Instructavatar: Text-guided emotion and motion control for avatar generation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.226236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.226236Z digest=sha256:df45e2b7e8e6d89824e2ba1d96bbdefd7d9ae77ae4d395ed639055d99e49fbff

Observation 8b556212-d124-413a-80fe-02527b13bd9c · outbound

This paper cites Human motion generation: A survey,.

Human Motion Video Generation: A Survey Human motion generation: A survey,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.287757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.287757Z digest=sha256:4c9c05bf3d8829d2c341ee07fd2abf1523fd5f6084b6a6fe92f6e9fcc6dc8d28

Observation 4f36bf34-f59a-48cb-9bbc-a03cf0aceac9 · outbound

This paper cites A survey of talking-head generation technology and its applications,.

Human Motion Video Generation: A Survey A survey of talking-head generation technology and its applications,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.350825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.350825Z digest=sha256:58a20e0c9b7725537ba9e0c2e3aec7a2034e4d3dfe42782814aa89d95257a313

Observation fe0bf50d-710e-4c45-bd0a-8ecb171f4c50 · outbound

This paper cites Unsupervised high-resolution portrait gaze correction and animation,.

Human Motion Video Generation: A Survey Unsupervised high-resolution portrait gaze correction and animation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.427366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.427366Z digest=sha256:3d474c17b9bf9c478d0f7796f1003612f10f2d5ceba16b59e657945df2ca089d

Observation 96978db2-617b-45c6-a88c-2b9dcd7adace · outbound

This paper cites Expression domain translation network for cross-domain head reenactment,.

Human Motion Video Generation: A Survey Expression domain translation network for cross-domain head reenactment,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.506521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.506521Z digest=sha256:75ae7cd2b5f9f895047a6b95a8c48c5705d81253112f76d8fdd0ef5e311bd207

Observation 15cb19b0-2feb-468d-917d-09e9e1f2921c · outbound

This paper cites Otavatar: One-shot talking face avatar with controllable tri-plane rendering,.

Human Motion Video Generation: A Survey Otavatar: One-shot talking face avatar with controllable tri-plane rendering,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.593467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.593467Z digest=sha256:b069e1873214acbc1f53513472a79fe3df477fe266b0946d9c2aafe8b8a93023

Observation 61e15e95-e4b3-4a44-9078-783033ffbf80 · outbound

This paper cites Follow-your-emoji: Fine-controllable and expressive freestyle portrait animation,.

Human Motion Video Generation: A Survey Follow-your-emoji: Fine-controllable and expressive freestyle portrait animation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.663654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.663654Z digest=sha256:810fe7456903608d09862d8c0401bdb8ef6312df04a004bcea94673de55efc91

Observation 32360012-a949-4738-80f6-3f8eadb82ceb · outbound

This paper cites LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control.

Human Motion Video Generation: A Survey LivePortrait: Efficient Portrait Animation with Stitching and Retargeting Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.724953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.724953Z digest=sha256:2c0307bbef3884bdb5eeec7eddd9165bfba39ff668a85c7c0d7d8d246dfc015d

Observation 56a2bb31-334d-4406-9232-280ad738af38 · outbound

This paper cites X-portrait: Expressive portrait animation with hierarchical motion attention,.

Human Motion Video Generation: A Survey X-portrait: Expressive portrait animation with hierarchical motion attention,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.800840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.800840Z digest=sha256:896981d42a1e505b75aa2a8291ec66666fc3fd3a2f4096ae8e51296b1868b1d8

Observation 83530f0c-6aad-4473-b9ad-45305848293f · outbound

This paper cites MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices.

Human Motion Video Generation: A Survey MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.872770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:50.885066Z digest=sha256:d5ab5b2a274f571774c67903bba31974c555d7efd7f3dd6ea5bfcff4dc20afc1

Observation c0ff62f4-6f49-463f-8945-49c096b8a056 · outbound

This paper cites Everybody dance now,.

Human Motion Video Generation: A Survey Everybody dance now,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:50.958740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:50.958740Z digest=sha256:d37e71e0d250f866220df1bcf867150c45dbd03e8b0b0dc3af9be2be23988034

Observation c22c8b7c-2ad8-42fb-9342-4e3d98fb9a9c · outbound

This paper cites Human motionformer: Transferring human motions with vision transformers,.

Human Motion Video Generation: A Survey Human motionformer: Transferring human motions with vision transformers,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.046502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.046502Z digest=sha256:0a07ca0edf0120143592c7e56d55d46e195f332f4ad66d65baa5a5a7a496c22a

Observation 36ec9f0e-890f-45dd-9bf3-6d155fb9f32c · outbound

This paper cites Bidirectional temporal diffusion model for temporally consistent human animation,.

Human Motion Video Generation: A Survey Bidirectional temporal diffusion model for temporally consistent human animation,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.122153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.122153Z digest=sha256:c768df5625891a2e4a3f8338a5459a5f67c0ffcfdb14bf04a14268f30c880edd

Observation 376dcc8e-92eb-48a6-8726-8ea2a0052aa4 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation,.

Human Motion Video Generation: A Survey Disco: Disentangled control for realistic human dance generation,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.184832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.184832Z digest=sha256:63dd4c81607cd4b7a5fe9f734685acc6f1b6f7c046b6ce5332a74548dda0a560

Observation 0e943829-ccb7-47e3-93fa-a4f5586821f7 · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation,.

Human Motion Video Generation: A Survey Animate anyone: Consistent and controllable image-to-video synthesis for character animation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.227878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.227878Z digest=sha256:50482e8ad0936e63a5e7e7d1302dddb229dc1be8e07e47b8769448f173080b63

Observation 3362df08-e56c-40f9-82cb-c6b2badd3684 · outbound

This paper cites Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling.

Human Motion Video Generation: A Survey Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.276195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.276195Z digest=sha256:e3df9e842a111c3484bd50e37800f464cd1983c808f732fe4fff7d5fe85eb22e

Observation fd85064b-470f-4b74-a4e5-6886a46b3cf6 · outbound

This paper cites Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer.

Human Motion Video Generation: A Survey Human4DiT: 360-degree Human Video Generation with 4D Diffusion Transformer

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.336374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.336374Z digest=sha256:22b0dbb458a8cfaf06f84ad69ea5f15a4817e0c0a4a5d3cfd8ede6d496df82ce

Observation 2b9aff82-d393-46fd-87cf-80804041afcd · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

Human Motion Video Generation: A Survey MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.408510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.408510Z digest=sha256:5c84c0925062aed57660d9a4f5d966af94bced09e29360a71af0d6f18682a25b

Observation c63e07aa-2336-4688-a64f-0b50089fa4a4 · outbound

This paper cites I2v-adapter: A general image-to-video adapter for diffusion models,.

Human Motion Video Generation: A Survey I2v-adapter: A general image-to-video adapter for diffusion models,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.491529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.491529Z digest=sha256:65cd966ca3c55479151173c45724c710a8611eb6cbdf381a2671b81615e73a49

Observation fa9be2f3-e5e9-410a-b530-6a312ebe37a2 · outbound

This paper cites ViViD: Video Virtual Try-on using Diffusion Models.

Human Motion Video Generation: A Survey ViViD: Video Virtual Try-on using Diffusion Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.546340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.546340Z digest=sha256:a502222e08c02f25f9dc755917254fca1df6ca9ce50391871aad70fc1782291c

Observation d2cc0be8-39bb-45d9-9c37-b4ecd5fd78aa · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion,.

Human Motion Video Generation: A Survey Dreampose: Fashion image-to-video synthesis via stable diffusion,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.631647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.631647Z digest=sha256:b2951618cb7c7286d9d36e84052400222db1e7e4290cb635205d869c89e4a314

Observation 9d8ee456-6b82-45e0-ad5d-8ea8e26f1ef1 · outbound

This paper cites Make-your-anchor: A diffusion-based 2d avatar generation frame- work,.

Human Motion Video Generation: A Survey Make-your-anchor: A diffusion-based 2d avatar generation frame- work,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.697233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.697233Z digest=sha256:8b790f1e9416ead44677e017171d7a48c078117938b043743f8fc0a5a9c372fc

Observation 427b93f9-4609-4dbb-9155-50f757a47e9d · outbound

This paper cites Write-a-speaker: Text-based emotional and rhythmic talking-head gen- eration,.

Human Motion Video Generation: A Survey Write-a-speaker: Text-based emotional and rhythmic talking-head gen- eration,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.768275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.768275Z digest=sha256:dc4d0ed5d221a08556b36f001dd83abd40f56e77bf236d8af1666c43f8f8c3cf

Observation 5039d190-f301-4c72-9d74-ef59776b9279 · outbound

This paper cites ID-Animator: Zero-Shot Identity-Preserving Human Video Generation.

Human Motion Video Generation: A Survey ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.833596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.833596Z digest=sha256:38a9d92ef3c3f0b086e576bb508aa7f60273bd29c37a09ab2829306c48bb0fb7

Observation 26db761e-aec1-4493-a7a7-cda9aa260e5c · outbound

This paper cites Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing.

Human Motion Video Generation: A Survey Edit-Your-Motion: Space-Time Diffusion Decoupling Learning for Video Motion Editing

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.888269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.888269Z digest=sha256:3bce2b5f0697f1fdb7a5ddae42bd4116f4fe61d4c8ea42d6280646d7fe1e6f15

Observation d555322f-e779-41da-9a18-576579a9c741 · outbound

This paper cites Follow your pose: Pose-guided text-to-video generation using pose- free videos,.

Human Motion Video Generation: A Survey Follow your pose: Pose-guided text-to-video generation using pose- free videos,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:51.996370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:51.996370Z digest=sha256:072ae887dd0ce087d13db000806228cca7b2518a7da1cf9a2e6d5d91cd002f63

Observation 9cc1b22d-2e2c-43b6-b574-ba33f0834aa2 · outbound

This paper cites Text2performer: Text-driven human video generation,.

Human Motion Video Generation: A Survey Text2performer: Text-driven human video generation,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.049764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.049764Z digest=sha256:8b5455cb291bb75510c0d76c7c69a386f9bef8267e72fde8f83dea7e20e8a583

Observation bbe3cd7c-9b7f-48d4-8566-34a95489be60 · outbound

This paper cites Styleheat: One-shot high-resolution editable talking face generation via pre-trained stylegan,.

Human Motion Video Generation: A Survey Styleheat: One-shot high-resolution editable talking face generation via pre-trained stylegan,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.111709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.111709Z digest=sha256:4022ee91074f98a4f1b9ba9804955a203329c410dd842710c3eca5402d08e043

Observation 6993dca4-834c-4cac-bfaf-32c063d5f9b6 · outbound

This paper cites Pose- controllable talking face generation by implicitly modularized audio- visual representation,.

Human Motion Video Generation: A Survey Pose- controllable talking face generation by implicitly modularized audio- visual representation,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.290746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.290746Z digest=sha256:7be8efc9096315f8371f625ac586f143112013ce58aaaacd27190d15973b47ec

Observation 09be9cb8-1621-43ec-9083-065b215e6636 · outbound

This paper cites Edtalk: Efficient disentanglement for emotional talking head synthesis,.

Human Motion Video Generation: A Survey Edtalk: Efficient disentanglement for emotional talking head synthesis,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.375519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.375519Z digest=sha256:d3e5a57d371fe8b644001de56e23e1ef343fd4e9a66c920d5cf0c466f23c69e3

Observation e47c006d-73d1-47e7-8e90-e6c2ecf9bc31 · outbound

This paper cites EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions.

Human Motion Video Generation: A Survey EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditions

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.470098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.470098Z digest=sha256:03ab875cf33f8884a54ee1b9eeeba08a95af7a820cdfa16ec2f6e5631b72a16c

Observation 908e1770-efa3-4e48-8a70-5ecf9d94803f · outbound

This paper cites Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions,.

Human Motion Video Generation: A Survey Emo: Emote portrait alive generating expressive portrait videos with audio2video diffusion model under weak conditions,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.561850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.561850Z digest=sha256:c633af969e99788120dd1a3320083015477f0a81d0ea35b8378e0c17714da4fa

Observation e0f06b7c-dffa-4f88-a39b-534522cb11ef · outbound

This paper cites Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation.

Human Motion Video Generation: A Survey Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.656632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.656632Z digest=sha256:a2cc99c1bf3c6c69ce688aca519fd16b2f7448b6df4ba7934f21e2a852e8c823

Observation 1840df13-0112-470c-926c-ea26659a34e1 · outbound

This paper cites Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose Generation.

Human Motion Video Generation: A Survey Emotional Conversation: Empowering Talking Faces with Cohesive Expression, Gaze and Pose Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.721044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.721044Z digest=sha256:698a56654865ab44172e32e4c7e18476107e60310a9a061a6bf263bcbf5a2dc9

Observation db688d56-7413-4c21-a9cb-29c65e141393 · outbound

This paper cites Makeittalk: Speaker-aware talking-head animation,.

Human Motion Video Generation: A Survey Makeittalk: Speaker-aware talking-head animation,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.772995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.772995Z digest=sha256:1f3ed5d5178168acec7ebcf57e386522c5fae368662eaaa51aef60fad987e8f0

Observation b815adc8-8abc-4055-8865-f5e35b3ae217 · outbound

This paper cites Live speech portraits: Real-time photore- alistic talking-head animation,.

Human Motion Video Generation: A Survey Live speech portraits: Real-time photore- alistic talking-head animation,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.824523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.824523Z digest=sha256:0761dc835323d1ff7a08968f98d9407f1d2e02bd6a8bc34064348e04041375d7

Observation 75aff3a9-89ae-4ee7-82fe-024be2d77f6e · outbound

This paper cites VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis.

Human Motion Video Generation: A Survey VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.868747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.868747Z digest=sha256:e4f50a7b8bd8135bbb5588999feff84caba50573dea0371fb057ab94ad81347b

Observation e4ed9920-2db5-4442-853b-631547e6c554 · outbound

This paper cites Dance Any Beat: Blending Beats with Visuals in Dance Video Generation.

Human Motion Video Generation: A Survey Dance Any Beat: Blending Beats with Visuals in Dance Video Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.680484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:52.930420Z digest=sha256:5b41a9a69273299fd6a68e3be37c20ac5fe74ed98745936fa02d7f9e8974c0d6

Observation 4dd2bf06-54a9-41ab-85b5-337b555a0042 · outbound

This paper cites Auto-Encoding Variational Bayes.

Human Motion Video Generation: A Survey Auto-Encoding Variational Bayes

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:52.986651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:52.986651Z digest=sha256:1b928c7fab7fd89013677b004918813b619a2f6a2509f4fb28a58dec86b1807c

Observation 3083795b-dd6a-49a6-b9bb-51d2e49b28e8 · outbound

This paper cites Geneface: Generalized and high-fidelity audio-driven 3d talking face synthesis,.

Human Motion Video Generation: A Survey Geneface: Generalized and high-fidelity audio-driven 3d talking face synthesis,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.070888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.070888Z digest=sha256:fdc8732ad9c1dee96edd45b0d3fb34fcdf09d97c6314b2f5aa0ca1a09b76e34a

Observation d79c4897-2ea4-4059-bd1f-e962667fb198 · outbound

This paper cites GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation.

Human Motion Video Generation: A Survey GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.123535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.123535Z digest=sha256:c692c890b5c4db89b5b918a69b536632eb05b717e2d224f259c58c8f6ccf864f

Observation cb5922ee-b2fb-48c4-8552-736c338f3a81 · outbound

This paper cites Neural discrete representation learning,.

Human Motion Video Generation: A Survey Neural discrete representation learning,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.184941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.184941Z digest=sha256:9a92437a3f69747f5e44ffe0d4105e2c8f7809013c65af19320f29bca65b9be5

Observation b3d7e218-04ec-42f8-a14d-a7c24694e47b · outbound

This paper cites Generative adversarial nets,.

Human Motion Video Generation: A Survey Generative adversarial nets,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.226243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.226243Z digest=sha256:897a7074447776f5c9f5a7f1022a43d35bfa340a2286a3b719868503a8c73958

Observation 8033ef9e-f150-46a2-af76-6b2d726a9db5 · outbound

This paper cites A style-based generator architecture for generative adversarial networks,.

Human Motion Video Generation: A Survey A style-based generator architecture for generative adversarial networks,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.303004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.303004Z digest=sha256:4ac880170aa005be37440b9acf526f3db2147e62f8058d9837f80ded8a3a8ab1

Observation c9f97cf7-0614-4fa7-989d-6b4aee2fb06d · outbound

This paper cites Analyzing and improving the image quality of stylegan,.

Human Motion Video Generation: A Survey Analyzing and improving the image quality of stylegan,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.394240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.394240Z digest=sha256:120a9b1bc61e7580e08f7a043f18775065cb961928fd7556f34ec103dd4131c4

Observation 85c06782-9c16-45a3-a75a-1911663f5447 · outbound

This paper cites An identity-preserved framework for human motion transfer,.

Human Motion Video Generation: A Survey An identity-preserved framework for human motion transfer,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.443233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.443233Z digest=sha256:4331559b2ae4cdb0094f93e71cadf29622333e651d24a13c7580ed8532467b98

Observation 39d5a18e-48cc-4112-b3e6-17f8765e922c · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

Human Motion Video Generation: A Survey Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.506844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.506844Z digest=sha256:39299eacaba55650b8fd3585f91df815dad84e57ffd25391d4047b79bb098169

Observation 48c3a0b4-1986-412c-bde6-5acd5cac82a3 · outbound

This paper cites Improved techniques for training score-based generative models,.

Human Motion Video Generation: A Survey Improved techniques for training score-based generative models,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.572864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.572864Z digest=sha256:0b005fa441584b3f88bd700582d6ed4d8b2f18e96ae03f51c4c791098f810c73

Observation e63960fd-1b14-47fe-b658-9c84020277ed · outbound

This paper cites Improved denoising diffusion proba- bilistic models,.

Human Motion Video Generation: A Survey Improved denoising diffusion proba- bilistic models,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.628440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.628440Z digest=sha256:b6aa2bbc92c51275bf7a61f5bd54c95b7ed1e3d6b70e9485cdb8b45f7d8139ed

Observation 41063b16-eb8b-410f-943a-b7554e3697c2 · outbound

This paper cites Denoising diffusion implicit mod- els,.

Human Motion Video Generation: A Survey Denoising diffusion implicit mod- els,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.693869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.693869Z digest=sha256:a6d99f3947ac82fd32bd613765565de311dbe0d076b8d409f24e190b91d9227a

Observation 23363c7a-46bb-418b-9d09-6ed5900097d3 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

Human Motion Video Generation: A Survey Diffusion models beat gans on image synthesis,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.755074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.755074Z digest=sha256:46f92e9ee79a9789982c1f2603fc23a13a390dc4ffd836273b9d12a78ea4e058

Observation cff4e806-623b-46ca-b538-4a87f6e4e325 · outbound

This paper cites A survey on generative diffusion models,.

Human Motion Video Generation: A Survey A survey on generative diffusion models,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.789291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.789291Z digest=sha256:dda786a624aeac66e5e2368b5f7f67daa6f739d7481d23c588f5589c931da890

Observation 902e7970-5ad5-45b6-ac03-5b2b8410360d · outbound

This paper cites Denoising diffusion probabilistic models,.

Human Motion Video Generation: A Survey Denoising diffusion probabilistic models,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.822594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.822594Z digest=sha256:6854e219c174073b9260dc159deeacb5dc647357da91fd978381495eff28bd7e

Observation 722e2ae5-ab12-4b38-97fe-1aa5f7795f8d · outbound

This paper cites Dance Your Latents: Consistent Dance Generation through Spatial-temporal Subspace Attention Guided by Motion Flow.

Human Motion Video Generation: A Survey Dance Your Latents: Consistent Dance Generation through Spatial-temporal Subspace Attention Guided by Motion Flow

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.622748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:53.851100Z digest=sha256:659f68e8ca96424e58a7d8c426bbe2c9893cff2f6453814f073a4e3fdddc6bd0

Observation 4ee471ec-042e-4a69-8fd4-a15731b82897 · outbound

This paper cites Human Modelling and Pose Estimation Overview.

Human Motion Video Generation: A Survey Human Modelling and Pose Estimation Overview

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.598466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:53.890018Z digest=sha256:6cc410b02f4fc467111d8919f002d7a4b244cdc7a08cba3a732037348c254090

Observation 1c81d839-0969-4e81-aed5-9c32b5da1054 · outbound

This paper cites Champ: Controllable and consistent human image animation with 3d parametric guidance,.

Human Motion Video Generation: A Survey Champ: Controllable and consistent human image animation with 3d parametric guidance,

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:53.932596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:53.932596Z digest=sha256:33396e7f483c4780853580f23aece08351cf1cce8002114347605b1af61aa911

Observation 3d3ca1b2-7a7e-41b2-a397-b9d0a406c75e · outbound

This paper cites Openpose: Realtime multi-person 2d pose estimation using part affinity fields,.

Human Motion Video Generation: A Survey Openpose: Realtime multi-person 2d pose estimation using part affinity fields,

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.005829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.005829Z digest=sha256:0a3aaad1020bc9befa272834b63ca51bca38eb16a6a3fe3be0e875f05b485ee4

Observation 00303c56-3dd2-43b6-a9c8-d02c764d3a83 · outbound

This paper cites Effective whole-body pose estimation with two-stages distillation,.

Human Motion Video Generation: A Survey Effective whole-body pose estimation with two-stages distillation,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.073205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.073205Z digest=sha256:e688b0238f3b934e181ff6a9375493c72866ec3985ae3d2516c0f7f8049a12ac

Observation fca9ccc0-861b-41cd-a53e-4046ced5ca35 · outbound

This paper cites VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation.

Human Motion Video Generation: A Survey VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.118365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.118365Z digest=sha256:d9fd4e59d0f8628857ad4cb2edfebc5d1a383bb9823efb2bcce702f8ae0f5be8

Observation 8cef5567-ea4f-42e2-8552-9e11eea741bb · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model,.

Human Motion Video Generation: A Survey Magicanimate: Temporally consistent human image animation using diffusion model,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.172725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.172725Z digest=sha256:7a44f880cbe011d18a57d0a010c33b3bb6abef1ba32e9809d43be6921bae7651

Observation c0371e81-1ba4-4557-89c8-b7d49adb864e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Human Motion Video Generation: A Survey Learning transferable visual models from natural language supervision,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.269247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.269247Z digest=sha256:bf304c9da9d218761cec8971b4855c033840bf6b2f9567b4cebc4644b45f1cef

Observation 53c38695-d68d-4e77-95e5-73b17a9f919f · outbound

This paper cites Conformer: Convolution- augmented transformer for speech recognition,.

Human Motion Video Generation: A Survey Conformer: Convolution- augmented transformer for speech recognition,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.305057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.305057Z digest=sha256:67aa459223306ed42b71bd07b033475c997c5364ccd78c5a4629bd1af3c6a79e

Observation 00de6092-d0b5-4c1e-b388-ef81c2fb428b · outbound

This paper cites Omniavatar: Geometry-guided controllable 3d head synthesis,.

Human Motion Video Generation: A Survey Omniavatar: Geometry-guided controllable 3d head synthesis,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.380434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.380434Z digest=sha256:cfd16f98561bca9d3d4dc4cfaded006cf0530b0836d9803195109008703fa108

Observation f3e72e65-8d87-413e-a9b3-442f8a2327e3 · outbound

This paper cites MegActor: Harness the Power of Raw Video for Vivid Portrait Animation.

Human Motion Video Generation: A Survey MegActor: Harness the Power of Raw Video for Vivid Portrait Animation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.430237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.430237Z digest=sha256:66bd6b66460d746d45d39992813ac0197414acbdb99e63aefc9e1d7913ab7ac8

Observation bb96ac18-cfcb-4dd1-addf-5fc27ca2a887 · outbound

This paper cites Faceoff: A video-to-video face swapping system,.

Human Motion Video Generation: A Survey Faceoff: A video-to-video face swapping system,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.497161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.497161Z digest=sha256:5962e2dee33e97612aa4851f92f340728e1aad5aa346390310d4d3e0a9caebef

Observation 1031deee-83e3-44be-ae43-5bebd3d0aa85 · outbound

This paper cites Finemogen:Fine- grained spatio-temporal motion generation and editing,.

Human Motion Video Generation: A Survey Finemogen:Fine- grained spatio-temporal motion generation and editing,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.633347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.633347Z digest=sha256:4931a85234ef9f276e93c1249355eb869752460ffab495bfb27bdfd94136e3fc

Observation adf6c3d4-e998-4b9e-b1a5-84e6f8448b0f · outbound

This paper cites Plan, Posture and Go: Towards Open-World Text-to-Motion Generation.

Human Motion Video Generation: A Survey Plan, Posture and Go: Towards Open-World Text-to-Motion Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.701074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.701074Z digest=sha256:b7086a167fb131a22f08e6c85d69e602120da754d80f77d0aabe90b44ce24c3e

Observation 4d30bd36-f327-41d8-b0e1-3e1bc4a9f982 · outbound

This paper cites Avatargpt: All-in-one framework for motion understanding planning generation and beyond,.

Human Motion Video Generation: A Survey Avatargpt: All-in-one framework for motion understanding planning generation and beyond,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.779686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.779686Z digest=sha256:c099bc4efe2b349247574db48949e5b0a93b1a3d56c49be95af1eaec91a2b5a9

Observation 1c3d735b-8c4f-4ff5-996e-a0d9bad80c5c · outbound

This paper cites Motiongpt:Finetunedllmsaregeneral-purpose motion generators,.

Human Motion Video Generation: A Survey Motiongpt:Finetunedllmsaregeneral-purpose motion generators,

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.842912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.842912Z digest=sha256:5067373502bc44345b84a15f6fa4d9724395c33808ea57015472f01a35caede8

Observation 1f47c7cd-23b3-45ad-89d5-e377bb05ad77 · outbound

This paper cites Motionscript: Natural language descriptions for expressive 3d human motions,.

Human Motion Video Generation: A Survey Motionscript: Natural language descriptions for expressive 3d human motions,

Reference 88

Resolution
verified exact
raw_fallback, observed 2026-08-05T10:36:58.521924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:54.889171Z digest=sha256:c947f09d7e825fdcf8e9400a86cb4135c2e1019f5ed71fa10d7e6b73e0c39b5c

Observation 52cd2ab0-f22e-4a3b-8213-f8f36c2ea07c · outbound

This paper cites Can language models learn to listen?,.

Human Motion Video Generation: A Survey Can language models learn to listen?,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.941478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.941478Z digest=sha256:3b0e8a16ebecb7651e52ee179f35990d6a0e57dc219a796bf48810aee0177b3c

Observation fae80bfd-7c96-4782-88fc-93930b81ffc6 · outbound

This paper cites Intercontrol: Zero-shot human interaction generation by controlling every joint,.

Human Motion Video Generation: A Survey Intercontrol: Zero-shot human interaction generation by controlling every joint,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:54.987297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:54.987297Z digest=sha256:af8324c39e17f5f097c563a63bf840266cbd1d2412193b12cc0a353754605b48

Observation ef22c8f3-2093-4d31-81c7-6e7635981c47 · outbound

This paper cites Digital life project: Autonomous 3d characters with social intelligence,.

Human Motion Video Generation: A Survey Digital life project: Autonomous 3d characters with social intelligence,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.061656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.061656Z digest=sha256:d8e2a08592e18f31cb5af8f6d2acdf801fb5c49fb499dc122402742784139bdb

Observation ab34fd6c-d07c-4b82-8c22-208514e94cc1 · outbound

This paper cites Style-Preserving Lip Sync via Audio-Aware Style Reference.

Human Motion Video Generation: A Survey Style-Preserving Lip Sync via Audio-Aware Style Reference

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.393948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:55.147557Z digest=sha256:8aed7907b0c22a11f50714d8f2fc84ba103cfc3cbbcfdc6b4e933e1b7d2c629c

Observation cef580a5-5bc0-4f25-9d36-dc9475c1b393 · outbound

This paper cites Dae-talker: High fidelity speech-driven talking face generation with diffusion autoencoder,.

Human Motion Video Generation: A Survey Dae-talker: High fidelity speech-driven talking face generation with diffusion autoencoder,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.219526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.219526Z digest=sha256:0ab66f77da5601425e900a1369fecf415bce6221b414ba62330a91145fd8f6ae

Observation 3940f10e-57e5-4fed-9b1f-410f2e027b65 · outbound

This paper cites High-fidelity generalized emotional talking face generation with multi-modal emotion space learning,.

Human Motion Video Generation: A Survey High-fidelity generalized emotional talking face generation with multi-modal emotion space learning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.341313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.341313Z digest=sha256:8903122e5fabef63650d2ed6e28cddf060fb5242e9a9415a56f0ac11e1fd9e80

Observation 6f19b458-6964-4118-b3e8-de5a03a081fd · outbound

This paper cites Do as i do: Pose guided human motion copy,.

Human Motion Video Generation: A Survey Do as i do: Pose guided human motion copy,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.437623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.437623Z digest=sha256:77cd1a80fe0a644ab8905edcda0b116e2a94f23854e7d94b79c60bad1d6b0e63

Observation a0884bf4-d635-46bb-87a5-e0fe8da07a45 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,.

Human Motion Video Generation: A Survey Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.499737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.499737Z digest=sha256:a92eeab86924779deabf08694c057e03f17dd0f7ddd6ae57479da0d8d5891a1f

Observation 0aa3248d-5c1b-43a6-8876-4d88a106f95d · outbound

This paper cites DreaMoving: A Human Video Generation Framework based on Diffusion Models.

Human Motion Video Generation: A Survey DreaMoving: A Human Video Generation Framework based on Diffusion Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.542297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.542297Z digest=sha256:515f4ee45e3d4982ba9f64d8170a0934c13ae7beac52c8c191959682709705be

Observation cb9c6960-cff5-4172-adef-5badaf5e1108 · outbound

This paper cites Disentangling Foreground and Background Motion for Enhanced Realism in Human Video Generation.

Human Motion Video Generation: A Survey Disentangling Foreground and Background Motion for Enhanced Realism in Human Video Generation

Reference 98

Resolution
verified exact
local_arxiv, observed 2026-08-05T10:36:58.352828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T10:36:55.595181Z digest=sha256:46aae6c9e6ea9b288a0621bc80f4a82ad1647ed3aff1fbf3daf93d38c1b0dfa1

Observation 6172c528-1170-46b3-a805-c2950208d190 · outbound

This paper cites MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion.

Human Motion Video Generation: A Survey MotionFollower: Editing Video Motion via Lightweight Score-Guided Diffusion

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.671513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.671513Z digest=sha256:b584a8832c4d72344f06d163c94e8561f26717fbd7e28259b2da0047e5e9a43e

Observation 67aa83bb-1e4d-451d-8d3c-178e608f0a1e · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

Human Motion Video Generation: A Survey UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-05T10:36:55.691493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:36:55.691493Z digest=sha256:580fedea35d7d35b176c40ca64d30acc54cb8c49821a1c06f8e6c238d58a9e06

Pith citing papers

No inbound Pith citation observations are available.