Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T15:09:02.727887Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 3 inbound Pith citation observations for arXiv:2604.11804.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T15:09:02.727887Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:58:52.344771Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-11T12:41:04.415342Z
79 of 79 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7c0e0ab3-8b5c-46a1-9ab6-317818d6db75 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation wav2vec 2.0: A framework for self- supervised learning of speech representations.Advances in neural information processing systems, 33:12449–12460
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0a11ef38-8f34-497f-bbe9-3778fef293b1 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Pyscenedetect: Python and opencv-based scene cut/transition detection program & library
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d78f62b0-a7ab-4b1e-a4f5-6106ce7fac0a · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c0aa2cf3-d9d9-4912-b478-6bb81d717bd9 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 69fa8af0-a40b-4691-8988-975eaecdbabb · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Posteromni: Generalized artistic poster creation via task distillation and unified reward feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d234a1bb-a95d-4a21-870b-5da508d07eaa · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a70e2c08-2720-43c8-ab53-4ee090b879ef · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e0deade-a0fc-441c-a69f-1d71fc7c4483 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Out of time: automated lip sync in the wild
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5dda2b6c-5a65-4c0b-b65d-ebe7ba1e6547 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Hallo4: High-fidelity dynamic portrait animation via direct preference optimization and temporal motion modulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b992870a-45a2-4560-9a3b-bdb6a46030d4 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Hallo3: Highly dynamic and realistic portrait image animation with video diffusion transformer
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2474ed55-67f8-4df7-83a7-40a46e12e8f2 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Cg-hoi: Contact-guided 3d human-object interaction generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2f250b84-4147-482e-b721-6df31f58e8e8 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Elevenlabs: The most realistic voice ai platform
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 43e4b1f9-9535-4e37-900f-e14a34faf8b9 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Scaling rectified flow transformers for high-resolution image synthesis
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8038e625-0a36-4e33-a68f-f695e43e8a77 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Re-hold: Video hand object interaction reenactment via adaptive layout-instructed diffusion model
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7be73ee3-1d8f-4aa0-a671-589d0b55a3ea · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Two-frame motion estimation based on polynomial expansion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d430c7f-fd44-475e-80bd-e20e4a92d472 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation SkyReels-A2: Compose Anything in Video Diffusion Transformers
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d78da58c-5a2f-425c-a00f-78ab574cdf1f · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HumanDiT: Pose-Guided Diffusion Transformer for Long-form Human Motion Video Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 12c4009b-f34c-4184-9261-dbdc18e49691 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5dde06dd-01e1-4411-83d8-9171573b25ad · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Nano banana
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 16312be4-5a15-43be-8551-a33f0c4f803c · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation LTX-2: Efficient Joint Audio-Visual Foundation Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c1d8d2a0-a7d1-48f9-9bb7-bff428615352 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6863fa26-4df1-46ec-b2ed-2a160aa76563 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Magicfight: Personalized martial arts combat video generation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation db41c143-0e15-4ef0-96ed-cddfe57c7860 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Dual-schedule inversion: Training-and tuning-free inversion for real image editing
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c8844a78-64e8-4506-aee1-dfe9f0ec77da · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation M4V: Multimodal Mamba for Efficient Text-to-Video Generation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 106168a0-6b39-4c67-a652-01b39ba9afa4 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5af15bff-4151-4342-8651-a88c127594c5 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Vbench: Comprehensive benchmark suite for video generative models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 85627e22-73e7-4684-b5a1-bea08e89bdd5 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b0e241ac-e508-4514-8400-ee23df8a4e38 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation GPT-4o System Card
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7646f2dc-3ccc-4346-b9bf-599f0d9694a9 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd328091-e011-475a-81cf-d0acb6ef7d7f · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Loopy: Taming Audio-Driven Portrait Avatar with Long-Term Motion Dependency
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f3b4905b-9c3c-4370-862c-3e9719185e14 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation VACE: All-in-One Video Creation and Editing
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d7b726b8-ea76-40d6-baae-3f7019829382 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Fulldit: Video generative foundation models with multimodal control via full attention
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 789d9386-e5fa-4bdc-9bb3-9bbcd33da32b · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Kling-Omni Technical Report
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8a522e79-6f7f-40c3-9027-d8ac000648e5 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation aedb031d-19e2-402c-90cd-b707372eda83 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b5d5222c-58c6-4e63-a675-b14c2f4d778a · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Omnihuman-1: Rethinking the scaling-up of one-stage conditioned human animation models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 224727c4-3375-483c-89c8-3fe684d959c7 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Apoavatar: Expressive audio-driven avatar generation via refocused audio-pose priors
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7d92686b-176f-4cef-9db2-f4464c4cb97b · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation JarvisEvo: Towards a self-evolving photo editing agent with synergistic editor-evaluator optimization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7db8cd8d-d960-41c3-a896-f37002162a86 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Mofu: Scale-aware modulation and fourier fusion for multi-subject video generation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6fca9cd2-8638-4bca-ac43-ed6b6e80d1c0 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Flow Matching for Generative Modeling
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d395414f-ee34-4f9b-b9c6-279fdbed163f · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Improving Video Generation with Human Feedback
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 78cd8348-c81b-4c6c-8519-b0c3160d0004 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Phantom: Subject-consistent video generation via cross-modal alignment
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 40f71188-02fd-412c-b9af-d41ddf073a70 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HiFi-Inpaint: Towards High-Fidelity Reference-Based Inpainting for Generating Detail-Preserving Human-Product Images
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 780d9e9b-b15e-475f-885f-ff553ae129c0 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Decoupled Weight Decay Regularization
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bd144a32-dee2-4f37-ac36-96724636dad6 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2b4c8b52-aec9-4833-8315-9f910e82535e · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Echomimicv2: Towards striking, simplified, and semi-body human animation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 81a642b5-b0b2-4134-a1c6-bdcf4de6de8c · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Hoi-diff: Text-driven synthesis of 3d human-object interactions using diffusion models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 57c6b94f-3749-46eb-a887-b090a8b5d12f · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Innoads-composer: Efficient condition composition for e-commerce poster generation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed4aee16-b160-4111-aed8-b9ac8270b8e5 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation MagicDistillation: Weak-to-Strong Video Distillation for Large-Scale Few-Step Synthesis
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 89f7c18e-ebbf-42a1-9d1a-8ae16292ae35 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HERO: Hierarchical Extrapolation and Refresh for Efficient World Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fbec8712-353d-4f58-bcd6-b1bc2dbf7421 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Scenedecorator: Towards scene-oriented story generation with scene planning and scene consistency
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5fe34f3c-5a62-48a1-8191-9cf55f247d50 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5396e65f-354a-42dd-89de-e8a75dad615d · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Ominicontrol: Minimal and universal control for diffusion transformer
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 79a901a0-f9ae-4edf-836e-c405c8ea8cc1 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Wan: Open and Advanced Large-Scale Video Generative Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation eb1f977d-5453-4bb6-a48c-b10c9c130485 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d2e36d4-16ea-45da-b32b-a61bb3019548 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Language model based text-to-audio generation: Anti-causally aligned collaborative residual transformers
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f9652e81-3c17-4d91-bb19-dae390acd5be · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0ea98015-ddf6-441c-bedd-44687b90c40e · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Fantasytalking: Realistic talking portrait generation via coherent motion synthesis
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2becd6ce-b95d-4dc3-b932-8384b7748564 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation In- terActHuman: Multi-concept human animation with layout- aligned audio conditions
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d741a0f5-5443-4135-8300-49bb3c0d2ecb · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Mocha: Towards movie-grade talking character generation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b9539934-7b00-4ba7-b8a8-243a28e219ed · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo 1.5 Technical Report
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1e47092f-291f-4559-800e-6c313b8ad0c4 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation D3D-HOI: Dynamic 3D Human-Object Interactions from Videos
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdb5a142-c123-418c-ad72-ad2b9f362a96 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Magicanimate: Temporally consistent human image animation using diffusion model
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e8344706-456b-48f0-bf70-7eb1f68e219e · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1de117d3-2386-4086-8726-75efef6dd041 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Follow-your-pose v2: Multiple-condition guided character image animation for stable pose control.arXiv e-prints, pages arXiv–2406
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ed8bc4fc-3e2a-4194-b89f-6285583a5ccd · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Hoi-swap: Swapping objects in videos with hand-object interaction awareness.Advances in Neural Information Processing Systems, 37:77132–77164
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 715c3ac9-04bc-4eb8-83ba-5f56206943a5 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Effective whole-body pose estimation with two-stages distillation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6d23955a-9035-441c-8a17-9f761317c7aa · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Diffusion-guided reconstruction of everyday hand-object interaction clips
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d85924bc-fd09-4bc1-82c1-58bee8dda5da · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation MagicInfinite: Generating Infinite Talking Videos with Your Words and Voice
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 47eeac02-5dbc-4b30-b9c4-37208fa942b8 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e641314b-9792-4f9e-b913-99cd95c05601 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Identity- preserving text-to-video generation by frequency decomposition
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2d4ab8e6-f913-4d1e-b7a5-a8b370588c47 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Adding conditional control to text-to-image diffusion models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 199cb090-1855-4413-888f-f4b95dbb5ac4 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Sadtalker: Learning realistic 3d motion coefficients for stylized audio-driven single image talking face animation
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f57c246d-ff62-4f4c-8b9f-9827500d268f · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Waver: Wave Your Way to Lifelike Video Generation
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1b6b86ba-0476-4aa6-afc4-89da3b0f88e7 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 07d5f765-e915-4aa0-9739-a987e41c1b6b · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Anytalker: Scaling multi- person talking video generation with interactivity refine- ment
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 366a78d4-111e-4e45-b269-d6f9b701451c · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation MagicTailor: Component-Controllable Personalization in Text-to-Image Diffusion Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6777b833-b724-4837-9248-097df5455720 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Identitystory: Taming your identity-preserving generator for human-centric story generation
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 333a1358-7d21-46e6-a6ae-6c04614f6027 · outbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Scaling zero-shot reference-to-video generation.arXiv preprint arXiv:2512.06905
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 588a0fcf-f17a-4baf-8d4d-98bc2e668000 · inbound
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b9e24e88-7ce4-47a6-b35e-5728b371dc09 · inbound
ReBind: Multi-Reference Video Editing via Structured Instructions with Explicit Reference Relationships OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adbe0d63-f762-49f6-80d5-4e0a436c0e7f · inbound
MoRoute: Dynamic Routing for In-Context Multimodal Video Generation OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.