Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T21:16:16.877500Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 16 inbound Pith citation observations for arXiv:2502.04847.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T21:16:16.877500Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:27:36.393512Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T16:27:09.332551Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0834b76e-0f24-4c5b-ae55-287b2553c5bc · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Conditional gan with discrimi- native filter generation for text-to-video synthesis
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2c869da1-5d43-4b8a-8e5e-ae3029a27fd7 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Multidiffusion: Fusing diffusion paths for controlled image generation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efeb2e9a-f12b-4b51-8b1b-bc4548846815 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Lumiere: A Space-Time Diffusion Model for Video Generation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e867c90-832c-4d76-b588-9e933755b13a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3d33765d-5839-4cce-859a-b6d478ef8e19 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f672520-901d-472b-8d6f-af74cccc060a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Realtime multi-person 2d pose estimation using part affinity fields
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c498ad1a-e33d-4853-abbe-c37987dc16f6 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Everybody dance now
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a7639469-4c26-4c8b-91b4-555d6c4612f2 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa6ddcd7-b2fa-40d7-b52b-ca67bd2e72b4 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Long video generation with time-agnostic vqgan and time- sensitive transformer
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eca8bde4-c838-4f29-a199-83c04fef5c19 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Learning individual styles of conversational gesture
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c20ca445-e2d8-46ea-8f40-bb6342833873 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Generative adversarial nets
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98ede29c-9ae8-4ec6-b8a0-7e943aa7e93e · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Talk-act: Enhance textural- awareness for 2d speaking avatar reenactment with diffusion model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 16d9b75e-2cd6-48fd-aa2c-d6c220f1d6c3 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3431fe6f-626a-4d1c-b21e-0696711b0326 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Deep residual learning for image recognition
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fc46ad49-c553-4da6-9e06-a105509aa82e · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea606338-560e-4248-b141-d20c78b9cf6a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Denoising dif- fusion probabilistic models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18a105c2-6925-4da7-8873-a1211c7f9486 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Image quality metrics: Psnr vs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 549d5fbc-a99c-45ff-8cbd-55a3a9cd6411 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Animate anyone: Consistent and controllable image- to-video synthesis for character animation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f203d0e7-d3a2-49ec-b015-f4888087e0f0 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Learning high fidelity depths of dressed humans by watching social media dance videos
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fbe126c7-0ba4-47d9-a498-0c300d1eed96 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Ultralyt- ics yolo
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b8c3bbd4-7563-4444-9d6f-4d9273549119 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation CoTracker: It is Better to Track Together
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caa24bc2-58d8-48b4-b717-e5cceb5bdbca · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sapiens: Foundation for Human Vision Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606a2cb4-fb2a-4153-a29e-930d68f22793 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Auto-Encoding Variational Bayes
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e654174-9035-441a-afb5-2289664958c6 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sequence Parallelism: Long Sequence Training from System Perspective
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38b59f59-1114-4173-b6b6-5c56b655cdb6 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Speech2video synthesis with 3d skeleton regularization and expressive body poses
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 95f0d104-63d7-4cfe-b340-b57b9e71ad3d · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb1ee01c-a96d-4f98-97e6-e1d5a1088a54 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio Motion Embedding and Diffusion Interpolation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89a2f01c-ce7b-454c-96ea-dbb90475d0cf · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Multi-task deep model with margin ranking loss for lung nodule analysis.IEEE transactions on medical imaging, 39(3):718–728, 2019
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c5a8a129-ed4e-4935-b30d-d81f3f6f7038 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Smpl: A skinned multi- person linear model
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52b64872-ff26-4240-b71f-97b34095901c · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Handrefiner: Refining malformed hands in generated images by diffusion-based conditional inpainting
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5aa0e527-979d-4b22-b76c-0616eaee2232 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Nerf: Representing scenes as neural radiance fields for view syn- thesis
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2496d139-34fc-46a5-b581-446bd63ead28 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Moore-animateanyone
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 233041cb-075e-4173-9bfa-5f6ccd144ef5 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Conditional image-to-video gener- ation with latent flow diffusion models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9f915bc5-4a7f-44ad-8975-4549ea24517f · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sora: Creating video from text
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 88ea8bda-db1c-48b2-aef0-7a104e4ea5a4 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Paddleocr
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation adfd2fdb-0242-4786-a0fa-7f47f3d2fe36 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Scalable diffusion models with transformers
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 09d26619-aab7-4ffd-8d26-9fda7e80d414 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Deep spatial transformation for pose-guided person image generation and animation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 24224dcf-4068-418c-8e1a-117086c20b20 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation High-resolution image syn- thesis with latent diffusion models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33fb1b53-ee0f-4e6e-80e6-e70dbbb0045c · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation U- net: Convolutional networks for biomedical image segmen- tation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1c91d50-98b4-44bd-86c0-0c9c894cd653 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Human4dit: 360-degree human video gen- eration with 4d diffusion transformer
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9dd8795f-b24a-43b9-9fe1-47d12aaaec3b · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Animating arbitrary objects via deep motion transfer
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 172ff92d-adf6-4301-98bc-71555fc6afc7 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Motion representations for ar- ticulated animation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2103c47e-0a16-4176-b17b-25aa5352f454 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d98cf96c-4cc7-44fa-92d9-a43133b90b35 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Denoising Diffusion Implicit Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e965526-5176-4704-94f5-9d875645fc68 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Roformer: Enhanced transformer with rotary position embedding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 782dfa92-6bcd-488e-b40c-904810c8543a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7baf9bc0-3533-4b64-9ab5-cce85c8a5e5e · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Animate-X: Universal Character Image Animation with Enhanced Motion Representation
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 373bfee0-0b59-44d6-aee2-a60a77e9bf8d · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9072b334-5323-40ea-89c2-7011a2b3a55d · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Attention is all you need
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9d46e001-4bc7-480c-8ff6-046dcf66f649 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Generating videos with scene dynamics
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dd1837a2-7651-4dd2-ab15-f431a74f2edb · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation DreamVideo: High-Fidelity Image-to-Video Generation with Image Retention and Text Guidance
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c50bc728-9176-4194-9a9e-54e4ec6930aa · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Disco: Disentangled control for realistic human dance generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2fb4275f-5ada-40e4-a547-0ca1db18da9e · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation One-shot free-view neural talking-head synthesis for video conferenc- ing
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 416fd5ef-2c35-41e3-8b4a-535d8e32da04 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e575bc99-aac1-4a4c-944b-0434b22fc388 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Image quality assessment: from error visibility to structural similarity
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24e9cf9d-bea3-4070-953d-868b2bae6971 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Hu- mannerf: Free-viewpoint rendering of moving people from monocular video
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2f038d34-90d9-498f-a2e1-6fd9ad51f590 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a55463b4-cb3a-4126-83c6-86b4632aa584 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Msr-vtt: A large video description dataset for bridging video and language
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation de0f70ff-e413-4a15-94a6-c9c1cc532d16 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Easyanimate: A high-performance long video generation method based on transformer architecture
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac3f24bd-effe-46ad-a0a0-0c2e50252347 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Magicanimate: Temporally consistent human image animation using diffusion model
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e3cb2976-eb72-4c37-9028-95cd3dc88840 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 555e26cb-a4ef-4315-815e-6292691ca29c · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Showmaker: Creating high-fidelity 2d human video via fine-grained diffusion mod- eling
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5fff86d7-cb6c-44db-a67d-26601bd47015 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Effec- tive whole-body pose estimation with two-stages distillation
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 18f37ecd-5158-45f5-9842-0e1ff2df2a33 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc77aae5-26e4-4c1b-b632-b2b5a52ef798 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation GeneFace: Generalized and High-Fidelity Audio-Driven 3D Talking Face Synthesis
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fad4119a-6fa2-4c5e-957a-98147c8bfe6d · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Make pixels dance: High- dynamic video generation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cd43f7a3-c811-407c-b541-57871757afc1 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Adding conditional control to text-to-image diffusion models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61ab8a81-7c20-4c62-86c6-16f684085ff2 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation The unreasonable effectiveness of deep features as a perceptual metric
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21d3a4cb-102b-424d-a8fc-06c0e4785f05 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 105b8940-7c80-4d26-a950-286a12b94711 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Tora: Trajectory-oriented Diffusion Transformer for Video Generation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 175a91c8-fe74-4340-9fdc-bbb350425cac · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation MagicVideo: Efficient Video Generation With Latent Diffusion Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 480f3328-3a0e-45b0-b6cf-a78765179e6a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation RealisDance: Equip controllable character animation with realistic hands
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16637732-4e53-428d-81ef-8d4ac8e1e2b4 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf959b04-8e84-4ff4-bc71-8419ed68adbf · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Sapiens [22] is employed to obtain pose keypoints, providing robust human pose detection for each frame
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bdb5b3d6-f602-403e-8425-878904bb3167 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation Unresolved cited work
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e3058091-ac8c-49d6-977b-1dfa8a8c378a · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation PaddleOCR [35] is used to identify and mark text regions in each frame, mitigating the potential interference of text artifacts with the generated data
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1259710e-9eec-40fe-a124-9612c5edb2d9 · outbound
HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation More Details for Data Collection A.1
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 07cca7ad-aab0-4be6-895e-9914d2213c04 · inbound
DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ac5ae84-74b1-4c0d-8aae-4c50a97ad4c1 · inbound
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c953dc74-b63c-44c1-9bdf-5a01e206b2bc · inbound
FramePrompt: In-context Controllable Animation with Zero Structural Changes HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 864de853-98ee-47f8-9c56-56039a9a2dc7 · inbound
CharacterShot: Controllable and Consistent 4D Character Animation HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2ed21f-c783-4b36-8408-d28c9f114b26 · inbound
InfinityHuman: Towards Long-Term Audio-Driven Human HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 230ce555-70fb-495c-a420-ec49c1992237 · inbound
One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 27613b73-29e8-4d03-b74f-6f1dc30001e1 · inbound
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 67e18852-6877-4312-8590-388f7765649b · inbound
AHOY! Animatable Humans under Occlusion from YouTube Videos with Gaussian Splatting and Video Diffusion Priors HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d78da58c-5a2f-425c-a00f-78ab574cdf1f · inbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a587e507-2fe7-47b8-a21b-f4ce67d6f45c · inbound
Reshoot-Anything: A Self-Supervised Model for In-the-Wild Video Reshooting HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6dfe205b-da8b-4496-a966-c1f14f0859a3 · inbound
EverAnimate: Minute-Scale Human Animation via Latent Flow Restoration HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5881a7e2-485c-4bac-a8df-f1ade38fdddb · inbound
Image-to-Video Diffusion: From Foundations to Open Frontiers HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7d27c7c8-caf6-4dfc-b1e9-bc628558a0d5 · inbound
Towards 3D-Aware Video Diffusion Models: Render-Free Human Motion Control with Mesh Tokenization HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2ac62087-02d9-4145-92ab-920eeae36bf1 · inbound
Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7dadeb01-d7c9-4a57-a229-86b334d4821b · inbound
Semantic-Aware, Physics-Informed, Geometry-Grounded Weather Video Synthesis HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 34bcd504-bca9-4d4c-92d5-282b38a8f4ba · inbound
ViDS: Video Diffusion Shader using 3D Face Tracking HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.