Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2505.04512.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T12:16:44.665393Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T21:10:09.669744Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8dd27455-0272-48ec-a382-50a76bb3bf6f · inbound
Hunyuan-Game: Industrial-grade Intelligent Game Creation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63e9d5a2-120e-4bf2-8687-2cb8e04ecf70 · inbound
OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abfbb6ad-57e0-46c4-8e59-66bedea81e69 · inbound
OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 788063b8-c336-4922-8764-25d8cbfc7f70 · inbound
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 074be566-ef03-4cca-86e2-8188ae16870a · inbound
DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fe3aad9-1022-4d98-a760-24ab29b51e31 · inbound
Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b641c452-348b-4b38-88ec-e82f018c0546 · inbound
A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a61444a-cf7f-48d6-94ed-a824c5c1e72d · inbound
DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 812b033c-b4c9-46d7-a44b-6d9dd7a06258 · inbound
InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5bf8773-d8bb-4ee5-9162-fd11c56f9371 · inbound
CustomX: Unified Character, Action, and Scene Customization in Video World Models HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ee46033-c95f-4400-813a-8b67dfa1477b · inbound
OmniCustom: Sync Audio-Video Customization Via Joint Audio-Video Generation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c97ce011-c67f-4cbf-a1d6-0403d5d9fecc · inbound
MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d24af3bf-84ac-43b0-ae30-3b1dd448ad90 · inbound
RefAlign: Representation Alignment for Reference-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4ce79ab-ccc8-420b-b23d-c466bd122994 · inbound
Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0163d288-73d9-4b31-97ba-08b701df773b · inbound
Evolution of Video Generative Foundations HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 204
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 04fe287e-dfd8-4469-bcbb-8a9e6bc57b04 · inbound
Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f08526f3-b1aa-49ac-a63c-5947baa19432 · inbound
Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1d8d2a0-a7d1-48f9-9bb7-bff428615352 · inbound
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 952199a1-ebad-4157-a3ce-83405553fdcc · inbound
Controllable Video Object Insertion via Multiview Priors HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 65333d83-84de-4bec-87d3-da71d12a5d45 · inbound
TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eacb0a72-0353-47e2-931e-c83cd4fbc94b · inbound
MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f3ac5160-9ad1-4dfb-b028-b606168c8028 · inbound
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d89586ac-4a70-4d6b-8094-c6a6994a90a8 · inbound
UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 586ba538-ab1b-4b26-8fe9-2a6b78fff3c1 · inbound
UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f54d741-72e8-4725-929e-dbc3f06fb308 · inbound
Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2af97305-875a-456f-9aab-b10078e1f81f · inbound
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a795f952-7525-4266-a9d2-855d4aa99514 · inbound
MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e690e724-3c34-452d-9f64-4b51f8002a07 · inbound
CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8d18100d-2885-40e1-892b-3bdca5f5cedb · inbound
HarmoView: Harmonizing Multi-View Constraints for Identity-Consistent Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0c8b23d7-dc6d-456c-9fb3-b8ba38c205fc · inbound
ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f39caec-ed05-4f26-aee3-ac525f169cd6 · inbound
DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf6267c4-094e-4ab3-9066-b21b8598b93f · inbound
Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd6a0b46-6194-4af3-baa6-6848e97a5470 · inbound
Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63374138-3231-45c3-9427-fbcf3959b7d1 · inbound
Keyframe-Anchored Identity Preservation for Sequential-Action Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18cd2add-2d16-4bc1-8d33-4e7a7b66130b · inbound
HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f34af647-a454-47b8-99ea-6fd7f628fb95 · inbound
UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7d9a41d-ccf7-4740-93f2-eabb73556743 · inbound
VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b68b13c-0de6-45ee-807e-dc1472fd2363 · inbound
Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.