Pith. sign in

Paper Citation Record · LEDGER

HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2505.04512.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.04512 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 41 of 41 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:14:36.019799Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T21:10:09.669744Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 51ea1a43-98db-4371-9d43-0f209bea8c12 · inbound

Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting cites this paper.

Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T22:10:42.704604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:10:42.704604Z digest=sha256:2560705cf6a90035efe3baa11175c192aa73ab1152902dcbdb4780a769f67e7d

Observation 8dd27455-0272-48ec-a382-50a76bb3bf6f · inbound

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model cites this paper.

Hunyuan-Game: Industrial-grade Intelligent Game Creation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.549695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:43:30.549695Z digest=sha256:977c2aa10456a24e5e7fee18eae4a5d36f981e5089be2749f9acc69c3b332b8a

Observation 63e9d5a2-120e-4bf2-8687-2cb8e04ecf70 · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:30.485928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:30.485928Z digest=sha256:7b222f4d371c5a8f45697f7214e5718db8cd51a82a766fd080a78acdd0f144ab

Observation abfbb6ad-57e0-46c4-8e59-66bedea81e69 · inbound

OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation cites this paper.

OmniV2V: Versatile Video Generation and Editing via Dynamic Content Manipulation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:39:51.666724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:39:51.666724Z digest=sha256:079000e630d81fa4deb3c29c419177466281bef197dd2c4267f24a87afa5f1ee

Observation 788063b8-c336-4922-8764-25d8cbfc7f70 · inbound

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement cites this paper.

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:32.934408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:32.934408Z digest=sha256:416f605088820f6f7dce39fb644c20aa8fff3869505226f29f382333ad073995

Observation 074be566-ef03-4cca-86e2-8188ae16870a · inbound

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers cites this paper.

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:37.363932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:27:37.363932Z digest=sha256:3d781d5807b2e3bb1f70b9e079ebb4e9a578611cefab4edf24ab4d28dae7655b

Observation 4fe3aad9-1022-4d98-a760-24ab29b51e31 · inbound

Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset cites this paper.

Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:09.055319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:09.055319Z digest=sha256:d080de6a42e5ad2f7ac5c9a10b75ae91609c18338d8f83df151b9d0c4f628c03

Observation b641c452-348b-4b38-88ec-e82f018c0546 · inbound

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality cites this paper.

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:26.356204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:26.356204Z digest=sha256:9a2bc5bd5febb093739c09f097f28dda7e12398ae0444140628e4779677c640a

Observation 0a61444a-cf7f-48d6-94ed-a824c5c1e72d · inbound

DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing cites this paper.

DreamSwapV: Mask-guided Subject Swapping for Any Customized Video Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T18:34:25.012542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:34:25.012542Z digest=sha256:021c419d19fcd1a1fa0dd88cc6c944381b66335fe15ce246e59103330e2b9827

Observation 812b033c-b4c9-46d7-a44b-6d9dd7a06258 · inbound

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion cites this paper.

InsertAnywhere: Geometrically Grounded and Optics-Aware Video Object Insertion HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T15:20:18.764867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:20:18.764867Z digest=sha256:2b9ebaff82b5dce51bb6c564a2026d7030d89f427b2c35e70c2bf389908411bb

Observation e5bf8773-d8bb-4ee5-9162-fd11c56f9371 · inbound

CustomX: Unified Character, Action, and Scene Customization in Video World Models cites this paper.

CustomX: Unified Character, Action, and Scene Customization in Video World Models HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T15:28:52.239638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:28:52.239638Z digest=sha256:5580dfc182a1bb6937246119d85289bc107b8f26b3a338503e4047e72927f58b

Observation 6ee46033-c95f-4400-813a-8b67dfa1477b · inbound

OmniCustom: Sync Audio-Video Customization Via Joint Audio-Video Generation Model cites this paper.

OmniCustom: Sync Audio-Video Customization Via Joint Audio-Video Generation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T00:11:12.422942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:11:12.422942Z digest=sha256:43836d7740688c7a6a430e0b3dd952a94a2e8f267209d32f483196ad84ca2d75

Observation c97ce011-c67f-4cbf-a1d6-0403d5d9fecc · inbound

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model cites this paper.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.169044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.169044Z digest=sha256:597f0c0b7ac8ffc9195f295f1efec2560d98f8d0fbc73b2be19cdf495bc2154e

Observation d24af3bf-84ac-43b0-ae30-3b1dd448ad90 · inbound

RefAlign: Representation Alignment for Reference-to-Video Generation cites this paper.

RefAlign: Representation Alignment for Reference-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-13T18:01:19.034578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:01:19.034578Z digest=sha256:5e8e9168e7e6ad676211f7b051686689e30d85afb5f15488ea3336e87ae3def1

Observation e4ce79ab-ccc8-420b-b23d-c466bd122994 · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:902b1e11a07fea4611fb9e2c33c5ec973b7cf3e03eb3c624e96d9e1042d51609

Observation 0163d288-73d9-4b31-97ba-08b701df773b · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 204

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.753975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:59107bf5c1f0b1d7008d3339696b8221d7a618ccd6576d28a1ba69102ec70d16

Observation 04fe287e-dfd8-4469-bcbb-8a9e6bc57b04 · inbound

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation cites this paper.

Prompt Relay: Inference-Time Temporal Control for Multi-Event Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:01:00.395392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T15:42:35.693519Z digest=sha256:17afb91ae39a0079e94e2c0bf6ae0c7ba1a7e855c7bc5d27bee3a2b81df3a0f7

Observation f08526f3-b1aa-49ac-a63c-5947baa19432 · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:03.662032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:9bac3d9f16c058d93de673530b2d82f771f33418d8341a0c750b2187908327d9

Observation c1d8d2a0-a7d1-48f9-9bb7-bff428615352 · inbound

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation cites this paper.

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:00.806138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T15:09:02.727887Z digest=sha256:0f4fded842199bd3615fa60753c4591f76691010a793c12a54f0a40e3a1fb2f0

Observation 952199a1-ebad-4157-a3ce-83405553fdcc · inbound

Controllable Video Object Insertion via Multi-View Priors cites this paper.

Controllable Video Object Insertion via Multi-View Priors HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:20.706056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:93c553ff30b28051bbf535f066015125f53115ca60ded265da1c5cf54a9752f2

Observation 65333d83-84de-4bec-87d3-da71d12a5d45 · inbound

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation cites this paper.

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:53:29.903427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T02:45:10.577070Z digest=sha256:c7239d4fd17cdab6c38e99f4453cdc23e1698fb8cd4f67a1c57d8b805fb0eb3f

Observation eacb0a72-0353-47e2-931e-c83cd4fbc94b · inbound

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation cites this paper.

MMControl: Unified Multi-Modal Control for Joint Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:46:28.018405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T02:55:09.008954Z digest=sha256:fe6c52b62e2d88734ce32ee3e55e9de1dba23aeb9661d7037dd738c6f0439e41

Observation f3ac5160-9ad1-4dfb-b028-b606168c8028 · inbound

FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation cites this paper.

FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:09.751609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T17:40:00.224358Z digest=sha256:e1afbc146f1e24db0459144094117a73519f90e5597438019d9dd4aa376cff36

Observation d89586ac-4a70-4d6b-8094-c6a6994a90a8 · inbound

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation cites this paper.

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:23.937729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T05:53:21.851578Z digest=sha256:e6ebb5e9af9c475961dd5da82e05c82feefffa2fb9de80b2d1f39b01858eba87

Observation 586ba538-ab1b-4b26-8fe9-2a6b78fff3c1 · inbound

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation cites this paper.

UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:03:03.463171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T22:00:01.349754Z digest=sha256:ae875585200193933a51a303768182d58797f5c634d214e0c188280966d83706

Observation 0f54d741-72e8-4725-929e-dbc3f06fb308 · inbound

Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation cites this paper.

Omni-Customizer: End-to-End MultiModal Customization for Joint Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:13:21.373265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T14:08:30.802619Z digest=sha256:e884809e9436a5321d1a9baca7714555dafa3475a0ea5cdbf3ed5e9d3ad72514

Observation 2af97305-875a-456f-9aab-b10078e1f81f · inbound

Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation cites this paper.

Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.501082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T15:25:22.778550Z digest=sha256:73185812ff9399e35b2d26162d6afdd9c5515069170df719a191b6bf81828c5a

Observation a795f952-7525-4266-a9d2-855d4aa99514 · inbound

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data cites this paper.

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.256912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T14:52:30.406683Z digest=sha256:a6503a5fd99a77eb1d89f197178fb5b623ffb1625176e703e665a6f3b1426ed8

Observation e690e724-3c34-452d-9f64-4b51f8002a07 · inbound

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation cites this paper.

CineDance: Towards Next-Generation Multi-Shot Long-Form Cinematic Audio-Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:07:28.122494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T17:30:25.371658Z digest=sha256:bf20755bd7a4d7a70a50a86cee73c6847dceba08c9d81095ffdba0f36d4caa84

Observation 8d18100d-2885-40e1-892b-3bdca5f5cedb · inbound

HarmoView: Harmonizing Multi-View Constraints for Identity-Consistent Video Generation cites this paper.

HarmoView: Harmonizing Multi-View Constraints for Identity-Consistent Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:37:36.656864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T13:49:50.272650Z digest=sha256:9f6ba7f2b3aecd0a9b729aa221addfba2e4aa54b681e279e9a951ebee9311459

Observation 0c8b23d7-dc6d-456c-9fb3-b8ba38c205fc · inbound

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation cites this paper.

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:17:45.961436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T10:46:56.871174Z digest=sha256:a164b86d992d60f936a89c1211c59edee566e690516c1b926add13ef11fb0bbe

Observation 7f39caec-ed05-4f26-aee3-ac525f169cd6 · inbound

DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation cites this paper.

DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-04T21:10:09.671686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-25T19:00:23.260939Z digest=sha256:480c54fdc44d914ed387f59d216b6cf25bb1e2c41b52130fd4a289c8cd1db82c

Observation bf6267c4-094e-4ab3-9066-b21b8598b93f · inbound

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models cites this paper.

Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:16:58.123698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-02T13:16:16.676244Z digest=sha256:fd3c5ee6f20d30074c03fafef15f5de31e4daf39e8cb89396a9496ead736d568

Observation fd6a0b46-6194-4af3-baa6-6848e97a5470 · inbound

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment cites this paper.

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-11T20:11:31.576642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:11:31.576642Z digest=sha256:50e1ef6fe2a62e546b40a6c0bd53debf27f3f576a526a602f4159b8f91844044

Observation 63374138-3231-45c3-9427-fbcf3959b7d1 · inbound

Keyframe-Anchored Identity Preservation for Sequential-Action Video Generation cites this paper.

Keyframe-Anchored Identity Preservation for Sequential-Action Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T16:30:01.157173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:30:01.157173Z digest=sha256:0307309a6bd20686fa3b64d7e125cae11b72a6fae484bdb18633231decd47ffc

Observation 18cd2add-2d16-4bc1-8d33-4e7a7b66130b · inbound

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement cites this paper.

HOMIE: Human-object Centric Video Personalization via Multimodal Intelligent Enhancement HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:20.695859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:20.695859Z digest=sha256:c3d200e0ee37a6cfca16d34e5977fc2d8a8d20923718fd2588dda791f9635d99

Observation 49884281-ad6b-42ba-a8a2-3f56dd32d611 · inbound

CoT-Edit: Let CoT Guide Instruction Video Editing cites this paper.

CoT-Edit: Let CoT Guide Instruction Video Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:14:36.019799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:14:36.019799Z digest=sha256:0bcdc211eb8c286b1525f00fb769bed224390ae1fc6d7cfed4a27b76b607c142

Observation f34af647-a454-47b8-99ea-6fd7f628fb95 · inbound

UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation cites this paper.

UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T17:55:19.705417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:55:19.705417Z digest=sha256:96955d847963a79623ef12d10f448ff6a02bd033d10abf6f4af6a154213df003

Observation f7d9a41d-ccf7-4740-93f2-eabb73556743 · inbound

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing cites this paper.

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T12:16:44.665393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:16:44.665393Z digest=sha256:db27172a4bf41749def5449d983d72bb4169808409aa926e9f676d71f88c4a5e

Observation 7b68b13c-0de6-45ee-807e-dc1472fd2363 · inbound

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation cites this paper.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.654243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.654243Z digest=sha256:1de5d395c6ea0960e428ddae8aa3140e455079f66c7bc1725688b2676772cf02

Observation ce2d5d91-1fd3-48c6-a8be-bdd31daa4ac5 · inbound

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains cites this paper.

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T05:13:22.764311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:13:22.764311Z digest=sha256:70e9d5828f5a904254820646b316e4ef5a83abfb49404b65cc5554846fad9d9e