Pith. sign in

Paper Citation Record · LEDGER

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

As of 6 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2603.14686.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.14686 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:51:18.368996Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:59:43.745654Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:09:56.485116Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b1aa708c-375e-4a52-867e-68f4f8ab4712 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.098812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.098812Z digest=sha256:bf51925dfb1264520b653821b3e914f9df1913dcc4a0690e17d0e5d5cc3c99b9

Observation e29c055a-350e-4ff8-9f65-44c65b6e6ee6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.105517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.105517Z digest=sha256:c94bb5480eee89fbc5893b8a5c9399508b7c0ee3f672a1e883dbcfdccf761768

Observation e55e8956-b8f0-4f28-8a80-3d82ee835944 · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.110781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.110781Z digest=sha256:8134c52128e9fa4fe9c688508dbbae9bfc774b5bc0cd44b5e6f5bcf3f9f22269

Observation b510ec62-daa0-4da5-ae4a-7cb0773c249d · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.115975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.115975Z digest=sha256:334d321bba3b45c65e8951033530ef703a35fc7e29e080fb467bbef05ca75d90

Observation d18c06b5-69be-4030-9fce-1223b936dc7b · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.120659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.120659Z digest=sha256:c24d3ffea7ba2f2a75ea223e50ae728fbc92ea9e30899f86570ef24050033e9d

Observation 9ff5ed00-7ae5-492b-b778-91a3e60a7ea9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.126833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.126833Z digest=sha256:60c21a5d0bbbbfbd197776ec5ee24c22b6eb2ddb1db4898cac5f62c0bd8e6faf

Observation 490134c9-613a-49b7-b5bf-be4e2c19bd48 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.138845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.138845Z digest=sha256:3e647c3610dff876800ac7dd418c427a631240a30adda47254bf7667296af175

Observation c6840a51-fe24-4d12-a38e-61f3f85afeb0 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.143944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.143944Z digest=sha256:b878355e195256f6bd7b2dfb1fecce8425506774c2cf87d736fa3e3c8b35c482

Observation 038ae1d7-c724-47d8-860d-9402ba8ba5e0 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.152357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.152357Z digest=sha256:b4a6cc53bd4808ffc389ec61e40d850a91e61c6ad2b2f7c2f0f32cad2aa520a0

Observation bb95b24b-9a80-490d-b510-3b6420cf8919 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.157830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.157830Z digest=sha256:36040236bee4543d019c527d6585c4d662a612a70d549df9d3f4dacb7afe30a2

Observation 133b439f-3eac-4fb4-9b53-14a50f0180d9 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.163164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.163164Z digest=sha256:91773bbc3245063e4eef0c2fa9745dda3932c20b65a707de4030120cedfe1080

Observation c97ce011-c67f-4cbf-a1d6-0403d5d9fecc · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.169044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.169044Z digest=sha256:d5946a858276222a9c4da816406c886f3dd0448f20ee8a0d7692d23365d74f10

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · outbound

This paper cites HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:c322947ad06cd801ce5bf22224534529ac7de33a016b7a68befc94612bac1299

Observation cd5008ab-a760-4c54-a698-5794c5011dfa · outbound

This paper cites RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.185135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.185135Z digest=sha256:67835063c29ed95f9ae2408bbf7304478fb0fe6480d31d54bb8b3d82b7a6d041

Observation ce23bc03-0137-4b26-82cf-d439a1628d84 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.192246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.192246Z digest=sha256:7365619769117dad6f191bb460c8cf8d9d3ec00e81313e50ea9933ab3a5eefff

Observation b218307f-9b56-46b9-9a4b-36196cf3f4be · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.204730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.204730Z digest=sha256:c99efc857f5ad646d3fc0f34a2de79a8df7d5f0a3748d1c75b1227a61810bd53

Observation c5b06e17-aa0b-4273-a73b-8b29735ecb1a · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Depth Anything 3: Recovering the Visual Space from Any Views

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.215614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.215614Z digest=sha256:2d78b2587d09d2f11883bd71de29675796b6fe8425c657acd791e2bf6482c809

Observation e21ede37-84ea-46d8-a57f-ad12acb28182 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.221154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.221154Z digest=sha256:551d075246eee5d389ae37ba70f7f4920260d7595ff0f16cda534862ad0e423d

Observation b16c105f-f6af-4215-ad0b-6fb297f91cdc · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Phantom: Subject-consistent video generation via cross-modal alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.232473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.232473Z digest=sha256:363e78617744ad1e2e348315f5f1807d20b31efbcf453ba0bf03ecc230f3daf1

Observation 16979f16-ebb9-4677-b32a-776376fd6c4f · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.237773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.237773Z digest=sha256:a05fc80fe915595adb1eaef9719085577cff018a091c976bb94197fe19151a83

Observation 04d54133-eb18-4228-a6f2-5658f12fa613 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.243683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.243683Z digest=sha256:e7812994342ee2113f9a264688ddddaa22116acaece077ce26a5ac90dd118186

Observation f01654d5-378e-4f14-9079-a98f9a64b87e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.255365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.255365Z digest=sha256:937bec156f6ae4ea6aacb0e209dd1be7fb23f8a153066271ce9ffbbe0da1140b

Observation 068d8b06-fb24-4022-888d-ab858b393888 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.260833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.260833Z digest=sha256:a0c042877092dd3dfc95e47e760c1bd66016e06ebab959c405dd1873c3650b74

Observation 5c40f80d-7768-49f7-a28a-f4b114bbe64c · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model SAM 2: Segment Anything in Images and Videos

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.267443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.267443Z digest=sha256:4ef4001ec12482c507482afe06bde9bc8bbfa315da4047fd8c0e9597b7b0c91a

Observation 04532f0f-0f71-4dac-8848-862446cf5bc6 · outbound

This paper cites InProceedings of the IEEE/CVF International Conference on Computer Vision.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF International Conference on Computer Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.248963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.248963Z digest=sha256:950ef4adb8ca0d1863647ecf84071902df101c9704b83e5d6f05784538f281cb

Observation 455c3f13-6a47-4e7c-af02-b4c4abf191af · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.278318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.278318Z digest=sha256:e14542031247e5e56d964fdc15faf9c73d762e82b396e35ea764b044b8256ad0

Observation 18119a37-0603-4057-9a07-671a26fcbc29 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.283800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.283800Z digest=sha256:2a3616ae056ed1aef37658ae2d3ddb1e430f25cf35413adcdfe93566238448fc

Observation e32f25a6-a742-480c-a4dc-8e7b60e0d4a6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.288763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.288763Z digest=sha256:770922865dc58f8402a3adee162d66f5a4b72dacc98682307ea6280b9176f77a

Observation 3f551690-9f63-4a02-b25b-03ae57310c35 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.272924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.272924Z digest=sha256:529a3214c5865e513df1d8481f5b6d27dd2af16df2d88d747c8323312ff55337

Observation 6aebffb1-0f1a-44a7-86ae-ab297ce6e65e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.300737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.300737Z digest=sha256:25456fbef2d8a74c2b414c05f44ffa125c976f9fde2d065860485b46d87863b0

Observation 6b6c40c3-0cb7-466a-804b-f41bee43aa1f · outbound

This paper cites $\pi^3$: Permutation-Equivariant Visual Geometry Learning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model $\pi^3$: Permutation-Equivariant Visual Geometry Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.305928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.305928Z digest=sha256:4f84233d68a02482c4b38e49b9cda89c2d51a3aa1cf9cb55943f90789e4a3bd5

Observation 03da2018-497a-4bf2-9c50-105062db989f · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.311084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.311084Z digest=sha256:068c24750bd2c0e572e89d9c0ab76214c2ad500c3281a1d49f7fcbb64c0d76cf

Observation be3e3eba-ad2f-4489-b69e-9ae609aadbb7 · outbound

This paper cites DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.294240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.294240Z digest=sha256:4b0920651532775a896ea025ac68ec220db266b7e997bd84148eae1825f36b7b

Observation 3040c207-a712-458a-84e4-e2076723cad9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.321386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.321386Z digest=sha256:2297e81faf56e3e8eea183eb4a7b24f7a0832be7bb04af9bc4136600aae2c85a

Observation 500fe1d5-d8c3-499e-a066-fcbdc3341779 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.326505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.326505Z digest=sha256:1ab91fb047391cb8a76bbc71a9bc63a880f6815738af9686fe3859465726b3b2

Observation c5a55f1f-e6f0-4b1e-bd12-009264df94b3 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.331988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.331988Z digest=sha256:bcaa72f1802dba7f0766ba87ba829618329be83cedd28afa152c892afa318c1b

Observation 6604432f-ebaa-4d44-a338-b4e6769e4fa7 · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.316887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.316887Z digest=sha256:0f9e72e518b8a4fadfc847e30f4d59e49e6015e1e3f9511f476997bb80f4579f

Observation a3c75d5f-5f17-46d8-a0a3-a045909e59d5 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.341928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.341928Z digest=sha256:fcfc121eec0d93c3f20038b6e309aa01ae99580171dc99962121148111af4ace

Observation 51ff2502-47e8-4c67-9e06-75716d625942 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.347090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.347090Z digest=sha256:2c20592b90e5b57a817a608e86f299b7636888f0f3a3ab54ce9c64698aee6a9f

Observation ac46fc77-f504-4418-8c2e-e76e0b9c3484 · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.351672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.351672Z digest=sha256:fdc2f330deba44bb3687d18d6cead84f20cad0d66562fe72cd794e6b994fcf7e

Observation ed2c303c-9e75-4f58-9952-0a77f0615b85 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.336435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.336435Z digest=sha256:08d4d47a76b6ea633a695f60c38897ff55189bd25d26254dd4ae14948a288b72

Observation 13bfb6fa-cde1-4de3-9fb3-a6e25a97abc5 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.362675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.362675Z digest=sha256:1d7a3d88979b5b4c860ef7ad71ce6cdde6c7e844045efd754d48a793c4b73d64

Observation bacbffde-6624-4b81-93d3-eaa7844ff749 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.368996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.368996Z digest=sha256:c508f4e7d641d56370982b81986b22f8dca337bb8f886ba99d332d105c37c22a

Observation 69f03735-c228-4fb3-b692-a0287f117ffa · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.357386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.357386Z digest=sha256:d34a8a6199ac448650ca42e459d089ae81718db9f5dca061eda08b59e5690a86

Observation 89f85e54-5606-424c-b4e4-237bd2bbec87 · outbound

This paper cites Flow Matching for Generative Modeling.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Matching for Generative Modeling

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.226041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.226041Z digest=sha256:944d5ad5e1c822c6bb07a9650a89de9890d03a46fbff381cb87b53e7246f4e1c

Observation ed24f181-dec8-4e7a-a322-aaee1267b8d0 · outbound

This paper cites InProceedings of the IEEE/CVF conference on computer vision and pattern recognition.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.132547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.132547Z digest=sha256:351a67ce0cd5a27d37b2967394eb58259b53ca00e37b57dcea961ba9313db557

Observation c19dee81-eaf7-4b91-be3f-d3db49139466 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model VACE: All-in-One Video Creation and Editing

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.200198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.200198Z digest=sha256:e3d89c49d82043ef27f61ffe4b30bd80d8b8ab66e93327aec3b2a858e956ac0e

Pith citing papers

Observation d380bd2e-4ecb-4882-b390-a636882722b2 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:524df733b6ba51cdf4ac4a93a6dfe50cc295414fffcad08f34dd55ec4f73b45b

Observation 38925e92-2ad6-4530-97bb-74311b3b8347 · inbound

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection cites this paper.

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T00:59:43.745654Z digest=sha256:2d5c6bf94652fce49386266376f373ee9e2b27469567a4a00e57c94852afa041