Pith. sign in

Paper Citation Record · LEDGER

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

As of 18 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 7 inbound Pith citation observations for arXiv:2506.08797.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08797 v3

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:15.192008Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:09:51.916002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T23:31:21.996122Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0ca87d1-4796-4256-ad53-f730ee11e621 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:14.990221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:14.990221Z digest=sha256:a6b1af35151e97d5c3f5367ddf0ae11533cb5d8993901b460bdb5f94b8bd3529

Observation ae459ba1-010d-4845-8394-1058f9fdecb7 · outbound

This paper cites Caron, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Caron, H

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.994895Z digest=sha256:166b17746c3630044a92fd2fe672d6f6622805e1d54d6a446985daf22a7965e5

Observation 0220a26c-8c11-4f9c-812a-d556b4e0d301 · outbound

This paper cites Chang, Y.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Chang, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.036510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.999349Z digest=sha256:01fb34d31eee677e734ff7d55dc1b9a9c5dca8eaef517cfef80decd63c72ff8d

Observation f97e95c5-d0b4-4ce1-91a3-0836e6316685 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.024565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.003301Z digest=sha256:8416f9c078c7f01e8091660e4ecfc906bbc6cacc1cb13969cf1b4fbc56b7aad2

Observation 4a7a9c56-3416-4d80-bd3f-af0a6ade30e2 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.012438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.007770Z digest=sha256:1b74861703a40505714ff000019838feab18da15506e3a3ef6532686f35aad25

Observation b1686fa9-01dc-4ba4-91c5-e233b32bcb2f · outbound

This paper cites VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.011660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.011660Z digest=sha256:c23f53ddd938895258d33e5f67822516ad978b4e39baf2f1b9277c65419af0f9

Observation ae9fa597-04f9-4565-bd5e-b4e03bb22212 · outbound

This paper cites Esser, S.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Esser, S

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.016076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.016076Z digest=sha256:51f8d227b1840777a3fb1672209c34a829ed5e2b595492d09d4008a3a8be4899

Observation de855897-b19b-43bd-b91f-bdbd81174438 · outbound

This paper cites Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:07:15.702938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.020125Z digest=sha256:9a791ce5f085f38cc4bd3625b6d76a7cedda50ff848030398230d2fadd0ed449

Observation 1bb3af1c-2a0f-4f5f-8cbe-f1d10a43893d · outbound

This paper cites Ghosh, R.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Ghosh, R

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.991435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.024580Z digest=sha256:83e1f5304cd2888ad3ef3f9522a706fcac9dbecd806b4582d8357fa660f725a4

Observation 66442418-fbac-4938-96ff-d3afea193c17 · outbound

This paper cites Heusel, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Heusel, H

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.980094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.028260Z digest=sha256:6cc80f5726dc2897eda25a2fd7ab5937baf24490b961b0916c8ec05da531211e

Observation 3eb2ed14-aaa6-4e03-952a-3acaaf760805 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.968197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.032053Z digest=sha256:0b1e1e67ac82a3e50a79586985fe60b766f28da8a100697f5b5d3c94f8796671

Observation 85e619ce-9d97-4571-8c6b-cad9d204c591 · outbound

This paper cites Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.035885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.035885Z digest=sha256:eb6dbe7a37b88064f88b0161f8e2113829adc31a2f47f6ea7114fe724042452d

Observation 65ac84c3-2cb5-4cf1-9e4c-e237b7ff3b63 · outbound

This paper cites Huang, F.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Huang, F

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.956274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.039605Z digest=sha256:3397c63c33f0eb5bca39850df6a7260be82f8821b489c6e302e635746b94b23e

Observation b60969b4-3581-4303-999a-c620a3b768f8 · outbound

This paper cites Sonic: Shifting Focus to Global Audio Perception in Portrait Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Sonic: Shifting Focus to Global Audio Perception in Portrait Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.043271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.043271Z digest=sha256:276c24dad6be021cf1de87309e716edfbb553bef49d9b082a5b751b1380dd4c4

Observation 09c495b4-da43-45ba-8c73-a0fa6310d9ba · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.944542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.047279Z digest=sha256:50eb77e209f95ff79e21ef885aee078e0df37018cb2e1e9f2be52411e126fb88

Observation fbc50b40-7c98-4fa7-864b-389ee098231c · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.932510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.050852Z digest=sha256:97d0d9db7d2c4c0d2df67dcdcd0add723eff458bc36f1a3de4b56386d54dbfe0

Observation bd07486d-a18f-48d6-ac7a-5c469db53935 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VACE: All-in-One Video Creation and Editing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.054368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.054368Z digest=sha256:946b926181e19b5b85aa96fbe2f586f4842acbe1336ccf7c569285f577e84ec6

Observation ee91f300-9ad8-4abf-8bb7-d6cea612440f · outbound

This paper cites Auto-Encoding Variational Bayes.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Auto-Encoding Variational Bayes

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.058186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.058186Z digest=sha256:d405e96a4f11a4c1c863b107c0517872a8c1bb6ca35abf243d891eeb6650c8f0

Observation 54dd8474-957c-45fe-8996-b5ec7fd95e48 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.061950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.061950Z digest=sha256:7d3bbd53e5a012af6a419951a27a82d8ba1de1cc7e8ec8d2156c16815a1e67b4

Observation ec9eb658-dbd3-48f9-a893-5ca6d3a0608a · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.920091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.065656Z digest=sha256:a614feab6fa4ab7c505ffd0a6b7ee7297c1fb5919769a73675cbcefe43090d6a

Observation cfd25029-0362-4163-8bca-e628478d122c · outbound

This paper cites CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.069517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.069517Z digest=sha256:ad2bf35a1c98045ea8693f995f5a6f5d4c0fffbc74b75094e159bb232f5cf0ca

Observation b23d7f39-b1a6-4e11-9b16-4b50e8928c45 · outbound

This paper cites OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.073459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.073459Z digest=sha256:de39a99c5712b1c83e28a0e62bae839e6d8816803e6c1d45ad876fb09c107486

Observation 0d936f60-8589-456f-872e-634135e14145 · outbound

This paper cites Flow Matching for Generative Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Flow Matching for Generative Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.077554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.077554Z digest=sha256:9951951eaa02badfc3c8c958e85a3e36dfc2cbe6a05cbab7c35d434be057bd85

Observation a5af6cab-0e8a-4acf-a7f5-623ac1baeb75 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.906640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.081439Z digest=sha256:28b1dd0b519defeb91a3d4f9e1afabb1dea68deb1a9d695b9b00b40c51c4181b

Observation 7d5c20ba-b5c1-4483-86d3-30b4e8471066 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.894523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.085543Z digest=sha256:414f85221c64151e13556872e5948826304b36bb12beaa409e266636bdfcab9f

Observation 617c4b12-378c-488b-b0ef-02e5e9dd2683 · outbound

This paper cites McFee, C.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation McFee, C

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.882539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.089605Z digest=sha256:3b58318e8b497199477ba073577b401cf780048a1f2a085a4ae9cf5c725c5db0

Observation e88431f2-74ff-4469-93e9-fb36edb08c7e · outbound

This paper cites MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.094291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.094291Z digest=sha256:505c47264efed7d554e99563ddd6dd68b89de0d1650ec84f1554664cfba13d44

Observation 65e7f9bc-1091-4f8d-a439-60a3744b5faa · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.098106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.098106Z digest=sha256:7fe0b9cb5b7f7ac1638ccd6adeb8ac0fea46dbb3ff995353c49a9f746e2c5fa7

Observation a4743763-47f5-4b1c-8470-485a02c0b32a · outbound

This paper cites ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.101686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.101686Z digest=sha256:06e2eb9250898ba74088760b3bc241539a0d33ffda8b80e9b4305e78868a6469

Observation 4dcea8d2-cb25-4ec2-9951-010cb855c598 · outbound

This paper cites HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.105677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.105677Z digest=sha256:940a910f408c91ea375a9fde9a84651da250b9bdd76e5b032eea2ec8c13dde3b

Observation 0ad89ee6-9fe6-4f38-a2cf-4f596193cb7c · outbound

This paper cites Radford, J.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Radford, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.870529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.109313Z digest=sha256:26d65f7e086a4a310fbe4a54c122a6ad6fdf831d3e8884fd78ab9ee8b1ced8c6

Observation 7bf6a82f-9aa0-4b4f-b799-aa2eb985cab8 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation SAM 2: Segment Anything in Images and Videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.112832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.112832Z digest=sha256:349f3da1a467fa29faa863f919a7e01bfb2c2904fbadb5fe0da58112f7ef9079

Observation 31de9a2b-2e04-436a-b098-dda4e6952cce · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.858624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.117225Z digest=sha256:4e1fcc6c1f056e70e47f3d0c2f2c4851b6a73b4cac1c4d149fde52e36f734dd4

Observation fd3f0e28-2d5d-45df-9678-7c1b18453a12 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.120920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.120920Z digest=sha256:cc5fd5a4a00619951f5fad4844899d2bda57a02042d18e0b11d67c94c153cca7

Observation 484fe3bb-2ce6-4ff8-8570-ae5d2a8f3313 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.838532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.124595Z digest=sha256:11f612491b980aa79e13123926476d56a33e911e1de6b9479a2908602b5a69e9

Observation a04d0853-748d-49a4-bd63-e9d7f45dcbd0 · outbound

This paper cites StableAnimator: High-Quality Identity-Preserving Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation StableAnimator: High-Quality Identity-Preserving Human Image Animation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.128129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.128129Z digest=sha256:ab19e2bc8113cd88d94d92f0266750dc4a960a76bddeaa43b5c14484f2f4abe9

Observation 6d1904a6-f7fe-4897-bd51-2227546b86b2 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.131757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.131757Z digest=sha256:571510aa4017e8f8df8ba25c1940c06fe491f2006f8cc7ccce78f0ef45eaf6fc

Observation d2dca16d-e51f-45be-82a3-418c67c2a856 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.135416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.135416Z digest=sha256:6b126330d5ceb1781cfcf2e83f61aa97c4a7f5e56297536720d53380d82cd042

Observation eae3a003-2f1b-4816-b084-79848605e6c1 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.139064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.139064Z digest=sha256:3266bda0f248aa514a546fa7bc74a8c4e3fc977eecb0036a412c12c27853a1a5

Observation 4838ac87-d887-4cd9-ab43-df53cbe08d6f · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.827007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.142718Z digest=sha256:f4ac71240397d1d27193f2210de4f8e99975e1864045202c33eb4b787f30349f

Observation 91ca9bd0-f838-4fa4-acdd-9ef2b09d78a0 · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.146443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.146443Z digest=sha256:4550590f6c2a136c4af7b84044bb20d9aa0376db6819915a1ce27846e5492976

Observation dc0a1690-b236-4020-bb5d-04958a9d41c7 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.150092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.150092Z digest=sha256:ea298932f57cf463dc5944f652aaee4f657c09857c3cea6a98a5417318c554e4

Observation 4f40aa06-c7c2-4967-9623-69976fed45b0 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.815546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.154172Z digest=sha256:5cefb5d738348fe5c6a8cc04428059491b435b383fe683af4d4414e82d47567e

Observation fd352863-e67c-4f0d-854d-90d3ee8b8360 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.803322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.157826Z digest=sha256:c4b648cf7cf76011aeaa175094183652b09f7caa5c5d661f0d9be0b8a2af817b

Observation d82759de-ba97-4831-9f09-2ce941a4833e · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.161431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.161431Z digest=sha256:d3948f457439fe7a7be74edd7ceaf5d0a0c5ae0e9444820820c576eab3d6a9fe

Observation 433e54dc-d2aa-4faf-8152-6f221b269c38 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.789992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.165245Z digest=sha256:ed987176df7ff7fbca2b8ed7f2bb6dc0966f5ef4d4997a91ef763dd9cb5ce9bf

Observation 84bebacb-0d3f-4e78-85fb-f212ed8f2ca5 · outbound

This paper cites HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.168770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.168770Z digest=sha256:2800522c35ac5ea0a84f4e27d77e1851ce4948927158fda5130e7e36d1862667

Observation 81cf7fda-803e-4bbb-9e3c-101e60cac3e8 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.778193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.172606Z digest=sha256:fd146f84ec232ec0eac62c3ee83cf342da51b066bf95e8ec0ce1d74446670228

Observation 8083b526-4c7a-4a20-a453-1c66d4e64b98 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.766481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.176525Z digest=sha256:7273879e8e996ad503f55d18dab252aaaa7f7576b8211a74c11da43a48846ad2

Observation 12113b3d-b607-4cd5-91dd-38834a3c3a54 · outbound

This paper cites Zhang, X.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Zhang, X

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.754349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.180425Z digest=sha256:8a9736195368eaaefe5ebc3c8a7abf9a3e0f25ed4b29c80dcdd6d95800c40e1d

Observation 81d175ab-acfe-476a-86f2-1be1d4a54a75 · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.184611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.184611Z digest=sha256:2f66e8519eefcb42f4a01d7247a0e952d94f847b7f745cb6ba3109e033663d7a

Observation 8298ced2-5661-40d2-a24b-03c834d10b85 · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.188299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.188299Z digest=sha256:065fc07a3d8ce5b390698555b7fcbaa1c00ea753e7f8815607b440458f38bcab

Observation 9891aa66-5325-4f7d-b603-e7a25078fd66 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.741878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.192008Z digest=sha256:452751d21b80cf56dc8ce2ef8376f16933ee03e4cda5e7d545efc67a77ec4bbd

Pith citing papers

Observation d4776e48-b2a1-4f33-9ac4-7b64aa05c48d · inbound

HOComp: Interaction-Aware Human-Object Composition cites this paper.

HOComp: Interaction-Aware Human-Object Composition HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:09:51.916002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:09:51.916002Z digest=sha256:4c1f4afe0e219f9c4b038fe9cef99694416f7499e78b1328e055d0728bc5cc9a

Observation 6520dd86-2b53-45dc-8825-bf9b92f07a76 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:dc16fe6cc64c90840d544a71e20dce2d7c4d51990f4474907f5ba4a6a8dddcb0

Observation 5cc0e35d-d677-4b28-ae37-737691ad3291 · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:a88e9bcdf37f3810e135cb1382e266a6eaa2abe90c73c66cf68c361e81e7c4c2

Observation ccc26b66-cc67-4ae2-a977-bf55d8a440c0 · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:4cccb87595e2ff250377543b47203788b028e78d8a806385871f3dcae8fde764

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · inbound

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model cites this paper.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:6f6418bb42b871fd6207e31fd7ad7785616da4a3cd97e02062b7b23839dfab2e

Observation 85627e22-73e7-4684-b5a1-bea08e89bdd5 · inbound

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation cites this paper.

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T15:09:02.727887Z digest=sha256:c9deb3578919a67ab2e7fe55b9a44a390f8fce5f34408964688995eeb4bde360

Observation 96207e89-b4cf-4340-a436-2da9dfdc30b4 · inbound

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation cites this paper.

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T10:37:53.829539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:37:53.829539Z digest=sha256:630418b808e9f28e659e27f33a17c8c6837598b63385248598a55efa590502fb