Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:45:10.247547Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 0 inbound Pith citation observations for arXiv:2412.04189.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T21:45:10.247547Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
82 of 82 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f642428d-8390-4af5-9e2e-5c12478cec54 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation The mug facial expression database
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3729c403-f6e6-4ebd-be85-0bfa87f6a5b4 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Detours for navigating instructional videos
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d056796-4471-4d98-8e3a-c9b3035531ba · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef9e9c3f-5e94-4e05-ae3d-8469909202c7 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Hrs-bench: Holistic, reliable and scalable benchmark for text-to-image models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 081a8a1c-63a7-44b8-916d-1cb2173a61ee · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation EditVal: Benchmarking Diffusion Based Text-Guided Image Editing Methods
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea11565d-4198-43ef-8e98-275e2cda8b11 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Brooks, A
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 31a2ed05-e1eb-4f98-a350-6f0915c6de41 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Gener- ating human motion in 3d scenes from text descriptions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 164c07f7-1778-4547-9ded-584d90d528d5 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Ceylan, C.-H
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 653a53e4-930a-455c-9b5c-2321bfceb89d · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Cognitive load theory and the format of instruction
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2446aa98-4e4f-453a-84e2-436b1a2a8a33 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Learning video-conditioned policies for unseen manipula- tion tasks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3bf0f642-0624-4c18-b885-de977ce05134 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Cheikh Youssef, A
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f1dc980f-22a6-483d-b76c-4f6416783538 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Videocrafter2: Overcoming data limitations for high-quality video diffusion models, 2024
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d422d952-13fb-484c-8976-fac8f8d821a2 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Panda-70m: Captioning 70m videos with multiple cross-modality teachers
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b93ae10d-05e5-4231-aa50-6a683b3c535d · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation AnimateAnything: Fine-Grained Open Domain Image Animation with Motion Guidance
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a078f0a4-2dce-4c56-95c3-d08364ef75ad · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Rescaling egocentric vision: Collection, pipeline and chal- lenges for epic-kitchens-100
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation af9eabd8-b54e-4094-8547-fa15793678a9 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Learning universal policies via text-guided video genera- tion
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 66797e2b-787c-43e0-ba36-820b7eeb3cd5 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Structure and content-guided video synthesis with diffusion models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c68ac1b6-f258-4cdc-a606-cdcb743c14bb · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Handrawer: Lever- aging spatial information to render realistic hands using a conditional diffusion model in single stage, 2025
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5be0a0e2-f208-4406-9f00-0b9e5df9ebd9 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Ego4d: Around the world in 3,000 hours of egocentric video
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 300d4442-7dcd-420d-9ed7-4b57719550eb · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Ego-exo4d: Understanding skilled human activity from first-and third-person perspectives
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c172d233-926b-4cbf-be3a-8421f482e350 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec3f5388-06bb-44f9-8e5e-8ffcae4ad229 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Denoising dif- fusion probabilistic models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c3099ee-0f63-4835-b28c-08df527f70b3 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Video dif- fusion models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a0aa761a-7ed0-40bf-9aec-66ccbc44ea45 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation T2i-compbench: A comprehensive bench- mark for open-world compositional text-to-image genera- tion
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5824e7d4-8696-4c6c-a8a4-8a053a525956 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Vbench: Comprehensive bench- mark suite for video generative models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8f37f44a-9c29-430d-85bf-2d5840cbbf56 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c817a730-3f91-4b55-8beb-f5f1a6508de5 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Vid2robot: End-to- end video-conditioned policy learning with cross-attention transformers, 2024
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation eb93ff5f-db30-46f1-9b59-6d6aebd881da · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation LEGO: Learning EGOcentric Action Frame Generation via Visual Instruction Tuning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 680d6f3b-8e02-4a57-b672-610afca8998b · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Temporal convolutional networks for ac- tion segmentation and detection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e39195db-2bee-4e88-8eb2-6b62179a1336 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Gradient-based learning applied to document recog- nition
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7b190191-8e12-4d35-ae9e-5f8621dff8e6 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Holis- tic evaluation of text-to-image models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca687851-383c-4a8d-984f-15c92b7ccb87 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f94d3f16-8793-46ab-8654-9082073dac07 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Egocentric video-language pretraining
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 98c6eb90-b37a-4244-8bf4-5efb8029c341 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d20e3a2-82c3-40c9-824d-5b1d9fec4651 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50bc3a62-629e-481a-8aa1-fbb9cc8b7afb · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Handrefiner: Refining malformed hands in generated images by diffusion-based conditional inpainting
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation df996d96-a8ef-47e7-a982-d9058f18d572 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation MediaPipe: A Framework for Building Perception Pipelines
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49015ec3-e604-4884-8243-cf92068f8b29 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Dexvip: Learning dexterous grasping with human hand pose priors from video
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bf88c734-6f0a-4a9f-aa4c-b6d1cf86dbce · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Recipe1M+: A Dataset for Learning Cross-Modal Embeddings for Cooking Recipes and Food Images
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86664d79-ab7f-40d2-b01b-9fddadd615b9 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Howto100m: Learning a text-video embedding by watching hundred million narrated video clips
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa91b37d-daa1-4561-938b-4830fd64619d · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Han- diffuser: Text-to-image generation with realistic hand ap- pearances
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0af6a45f-b2e8-4f19-9e17-be6b9a04bc55 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Conditional image-to-video gener- ation with latent flow diffusion models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f6768da8-c1d6-4744-bbc8-cf575ce84005 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Ti2v-zero: Zero-shot image condition- ing for text-to-video diffusion models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 78ee2f4e-1f22-4847-8472-d2e9bcb3e993 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Improved denoising diffusion probabilistic models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4102e6b-ab42-4357-a7fb-06d1fcc38034 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Learning transferable visual models from natural language supervi- sion
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e52dd587-8d4e-46dc-b45b-69cbee120526 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation The meccano dataset: Understanding human-object interactions from egocentric videos in an industrial-like domain
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef77aa70-7a2d-415a-9c0f-a102233f68e6 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation High-resolution image synthesis with latent diffusion models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bc4e86e4-e2c6-4ed0-89e9-ecd873b53b03 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Photorealistic text-to-image diffusion models with deep language understanding
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c642cb2-8f57-4ffd-9513-3129aa1ebdb9 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53feb2f4-ee4d-4333-943d-3dbf01ed72a6 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db523071-473f-4964-9366-d7e124177089 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4df0f93-d0e2-4c0e-b988-69f5664eb36d · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Many turn to youtube for children’s content, news, how-to lessons.Pew Research Center, 7, 2018
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 66cdd9d8-2910-4151-a238-2f0e5b04c7f3 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Denoising Diffusion Implicit Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a897df8d-f6e3-4e7a-8dfe-13e88d87d733 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation VideoAgent: Self-Improving Video Generation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c95c5cdd-8d6b-448d-83a9-3522f8228f47 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Genhowto: Learning to generate actions and state transformations from instructional videos
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c43752f6-ee06-4057-b6f4-c3d3a596b5f7 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Fvd: A new metric for video generation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1b24bb18-76bf-4fe2-9906-07aeb04a19af · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Long-term temporal convolutions for action recognition
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 31ee753b-b15c-464d-88bb-54425c944b60 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Attention is all you need
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b87055f-1e6c-4f93-8fb3-8744c03b4455 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Imagen editor and editbench: Advancing and evaluating text-guided im- age inpainting
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57f8ec96-e070-4281-9377-522da41f2d4e · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world supple- mentary material
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 31f1c55b-a807-4f1e-80fe-b475490bfc37 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Multi- modal augmented-reality assembly guidance based on bare- hand interface
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 01d2e141-519f-4a36-bfba-906867de73c2 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Holoassist: an egocen- tric human interaction dataset for interactive ai assistants in the real world
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2f5d3fa-986b-484a-b3ef-a2dd0c6c8b6b · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Videocomposer: Compositional video synthesis with motion controllability
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f5f1db1-2755-4779-a441-58553b189d1e · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95626d99-19c3-4dac-a75d-04c1eacc7e6c · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Lavie: High-quality video generation with cascaded latent diffusion models
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c3a86d49-5613-41b7-80cf-553565c35d36 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Towards A Better Metric for Text-to-Video Generation
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2d1a377-f961-4b2d-9d36-2c0252e97de2 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Freeinit: Bridging initialization gap in video dif- fusion models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a319beee-3af4-4f41-ba47-b358f3743d64 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 69cbdec1-6388-4ac8-b1f9-b99551f26838 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation X-gen: Ego-centric video prediction by watching exo-centric videos
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1ed1cdca-47e3-4de5-aa48-68d8ba40aacb · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Ad- vancing high-resolution video-language representation with large-scale video transcriptions
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f616454d-f686-4b37-812b-d70680323439 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Stat: Spatial-temporal attention mechanism for video cap- 11 tioning
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7fe80857-1b75-4ccd-bb03-8787d860ba96 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Learning Interactive Real-World Simulators
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4d2ce0d-5e36-412d-9e67-848dc9fd2a39 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Annotated Hands for Generative Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3b0d8d95-bca7-4ccd-9b57-b725e623dc4f · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9af77c-8673-4a21-9319-b37c2a7c5d11 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Learning universal policies via text-guided video generation
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 833fe562-249e-434f-8fb5-ff90f3d117ad · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34151f3a-bd05-4fbf-9a39-1978b2539da7 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1800aa40-4ab1-4cab-85ee-30f1711d8c9e · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation def0720d-1a4a-4ebb-835e-c95ad5a8c746 · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Pia: Your personalized image animator via plug-and-play modules in text-to-image models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 503325e1-b18b-4b60-b170-f8db274132fb · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Open-sora: Democratizing efficient video production for all, 2024
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 180a3c7e-14d0-45b2-bee9-7b78ef2a8c7f · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Towards automatic learning of procedures from web instructional videos
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e119dd37-e116-46a1-90e5-68fbe57e7a6b · outbound
HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Motion Control for Enhanced Complex Action Video Generation
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.