Pith. sign in

Paper Citation Record · LEDGER

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

As of 21 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 5 inbound Pith citation observations for arXiv:2506.02444.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02444 v3

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:06.878705Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:10:14.672688Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T00:18:30.107372Z

Reference resolution

72 of 72 outbound references displayed

  • verified exact1
  • verified fuzzy30
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fc5e5238-9f2a-4b5e-ad10-7e160adb5df5 · outbound

This paper cites Body of Her: A Preliminary Study on End-to-End Humanoid Agent.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Body of Her: A Preliminary Study on End-to-End Humanoid Agent

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.252913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.252913Z digest=sha256:8b4ad4af044975ea03c0e78fb801e955c51fe0cc4ad192c43885734697c2931a

Observation f7f07dd0-afa8-4423-9915-89916cf3086f · outbound

This paper cites Qwen2.5-VL Technical Report.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.438395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.438395Z digest=sha256:93813e200cc91d2b3f6aa68826d587f299788bf34bc28252d27415c355f8ef69

Observation f7cb1d75-dffb-4694-8c49-a23ef5ff2ab0 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.510279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.510279Z digest=sha256:aa0a836bb556af8977a54ea02a0a8989db4625e4146b7142d100f6a4e8f29893

Observation 7ddaac18-9d8d-4a40-85d2-b7c3891e1941 · outbound

This paper cites Physically plausible full-body hand-object interaction synthesis.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Physically plausible full-body hand-object interaction synthesis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:15.278550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:28:59.562898Z digest=sha256:ef01221841a39eb07ed3bb3851e1df30a76788548f2ecaba260d6850709c0022

Observation 88f12b54-76ff-4f82-99a5-afcfbbf7ba93 · outbound

This paper cites Video generation models as world simulators.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Video generation models as world simulators

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.634748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.634748Z digest=sha256:5af3113cb64859eb2455221321be6bb0e8f3ad4f27c62ccb66babad2169bf242

Observation 376c7fd6-a8d1-4fe7-8652-8b4a68d62048 · outbound

This paper cites Emerging properties in self-supervised vision transformers.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Emerging properties in self-supervised vision transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.770914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.770914Z digest=sha256:59e0aa29f8b022b75afca7ccda8b1f2795532f673d6e156f533f7f816e1b5fee

Observation 6068e9ad-7fe3-423c-b148-398a55b5d045 · outbound

This paper cites Text2hoi: Text-guided 3d motion generation for hand-object interaction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Text2hoi: Text-guided 3d motion generation for hand-object interaction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.890753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.890753Z digest=sha256:18c06e8c1f64e96637f8f402d7ca98f251c18f34793b935586c3f1c2e43e2a00

Observation 1499def3-c35a-4461-80af-35ccc4d7ead7 · outbound

This paper cites Dexycb: A benchmark for capturing hand grasping of objects.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Dexycb: A benchmark for capturing hand grasping of objects

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:28:59.947785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:28:59.947785Z digest=sha256:b84114452c1e430c291706ada07d10e772fe319a5c45096596683af320175442

Observation 3c290931-61bb-4ed0-853e-ddf3935892c2 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.054337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.054337Z digest=sha256:253b85e00e535b2ce8ef96321bf6403fd749a055eb356bfabe180514e6f5b3cb

Observation 0e3d241b-db13-453f-a9ba-7ba44c20972f · outbound

This paper cites Diverse human motion prediction via gumbel-softmax sampling from an auxiliary space.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Diverse human motion prediction via gumbel-softmax sampling from an auxiliary space

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:14.906162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:00.205296Z digest=sha256:4b222089182387eaaaab3aea28e619c5a7c65b4c9d6f762b96dea4ecdc8b4ba0

Observation e6b4d1d3-6fa3-4107-ad08-2a7ae6231d7e · outbound

This paper cites Cg-hoi: Contact-guided 3d human-object interaction generation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Cg-hoi: Contact-guided 3d human-object interaction generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:14.670755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:00.320954Z digest=sha256:64fd80746a39942e067c0ea3c7a514f4ce57975afc91f9c967cfa10033423780

Observation 6bbf3627-d530-4fca-b477-e617644b11f4 · outbound

This paper cites Arctic: A dataset for dexterous bimanual hand-object manipulation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Arctic: A dataset for dexterous bimanual hand-object manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.418663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.418663Z digest=sha256:ec3aa91cacfc4d1662a58bebb2389367c6f38aaffe88a5a44afe9990295ca816

Observation b7e3a644-57d3-451e-9ec5-c6e17944cbb1 · outbound

This paper cites Coohoi: Learning cooperative human-object interaction with manipulated object dynamics.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Coohoi: Learning cooperative human-object interaction with manipulated object dynamics

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:14.456528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:00.565115Z digest=sha256:64a1bcb4cf940c3414ffcc57f6bf097ebc9c1d4dfc217484d3de2b412687cde0

Observation 15868f4e-08f4-4729-b0e6-44e50208efcf · outbound

This paper cites Prediction with action: Visual policy learning via joint denoising process.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Prediction with action: Visual policy learning via joint denoising process

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:14.221492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:00.671840Z digest=sha256:cc43e44901f7ae36e34c6451a7b96ecbb87640762081fff8c2d7da0d5d863bae

Observation a43d552e-a8ea-46b4-8de5-27ee47bc0eaf · outbound

This paper cites Stochastic scene-aware motion prediction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Stochastic scene-aware motion prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.783005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.783005Z digest=sha256:15b84b2b9f8515a0229176d1041e986a4dd2951caa18d2e0ae6c66f290de9fbd

Observation 90dfb212-eea2-47ec-8c03-fd8cfe0030ff · outbound

This paper cites Denoising diffusion probabilistic models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Denoising diffusion probabilistic models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:00.916542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:00.916542Z digest=sha256:366e55d20e69da54171a7ac4957c253054cc9d376a9c3e373ab20fb87d5bd596

Observation d37851e6-0b83-45df-92d5-90ef879c4e3b · outbound

This paper cites Cogvideo: Large-scale pretraining for text-to-video generation via transformers.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Cogvideo: Large-scale pretraining for text-to-video generation via transformers

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:13.932132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:01.016743Z digest=sha256:5dfb3ad8c3b20f0e4b49ad3a59821f7d14ade6a29d5833cfbca56f0bf7abcf9a

Observation af3ea639-e121-414f-a1c8-8489ba011bc3 · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:01.156559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:01.156559Z digest=sha256:1ce57fa4e722de9cd45202c2f017041ae1e85fb69660cbcb0367f2745034e07d

Observation a3d023e0-872a-472c-af94-869ef263f45a · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Vbench: Comprehensive benchmark suite for video generative models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:01.232642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:01.232642Z digest=sha256:8c3920e12e75552d73e5fe47fa75d7c2b521bdde5bb71cc5bbaa213a2ef96f7c

Observation 1f83025c-a539-4be8-b4d9-e1537f72974d · outbound

This paper cites The power of sound (tpos): Audio reactive video generation with stable diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios The power of sound (tpos): Audio reactive video generation with stable diffusion

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:13.679724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:01.343728Z digest=sha256:e0da481e1af1b30b57a8ffb5bdce94c656627ae19decc3aec0611445d9a625f7

Observation 862781df-f76c-4a6c-8594-e3c06f35902c · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios VACE: All-in-One Video Creation and Editing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:01.468907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:01.468907Z digest=sha256:6bfa18bf2ba8ed1bcbe27e2c749a81ca3f6a06a6db4e240ad68def726cfee516

Observation f1fa39dd-4b4b-4d98-aba5-f0d28ec6b820 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:01.608522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:01.608522Z digest=sha256:29408456bbc0594172e8ab990fb055cc8f0a4de32a953a1764adac320742d8d8

Observation 6509e18e-422a-40d1-978e-9d2b3bddf12f · outbound

This paper cites Nifty: Neural object interaction fields for guided human motion synthesis.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Nifty: Neural object interaction fields for guided human motion synthesis

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:13.461294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:01.712935Z digest=sha256:5ef5435ef97b134855a3652de0f7e6a6d4a034e85a5ac077cc82fb12c9c948db

Observation 704b2776-f350-4a1e-8336-33cfe9d43e40 · outbound

This paper cites Interhandgen: Two-hand interaction generation via cascaded reverse diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Interhandgen: Two-hand interaction generation via cascaded reverse diffusion

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:13.229870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:01.841671Z digest=sha256:42bf2a1102bdcf87e54e516440bc4c9eaf34b91ac161dc9bac68a26e3ad6ae71

Observation 6061c08c-5794-4aa5-b5ad-23aa0503fb3b · outbound

This paper cites Controllable human-object interaction synthesis.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Controllable human-object interaction synthesis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:01.973487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:01.973487Z digest=sha256:020b012c997949ee204253eabd402a6c85f7c33a267383c4c1f9fee237f4669a

Observation 4ad78e16-7d2f-42f1-8162-63c01103766e · outbound

This paper cites Object motion guided human motion synthesis.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Object motion guided human motion synthesis

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:02.099180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:02.099180Z digest=sha256:89ca04a16ff6fdb06e4d56979240c28c82ecd92eff91042169cc814a02cddebe

Observation 1a37028e-a14d-45fc-a6c3-28b8ecf7f0fd · outbound

This paper cites Task-oriented human-object interactions generation with implicit neural representations.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Task-oriented human-object interactions generation with implicit neural representations

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:13.077736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:02.159259Z digest=sha256:270a6854f66cf2b7e9d8c450c84348aa0d9aeaf902bde4942239075bd3b1af29

Observation 778cd270-11c7-4e00-9a06-cdefb38374af · outbound

This paper cites Vision-language foundation models as effective robot imitators.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Vision-language foundation models as effective robot imitators

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:12.879061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:02.272037Z digest=sha256:313854981e0ff0b4a3388bc2106bf9bbbe6ae82fb0479cb947f5d278ad4d103b

Observation 7841240a-2b58-4190-a8ce-a0cebf37014c · outbound

This paper cites Amt: All-pairs multi-field transforms for efficient frame interpolation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Amt: All-pairs multi-field transforms for efficient frame interpolation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:12.671992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:02.412088Z digest=sha256:5c4d4a54179461afaadc5e8e95678691dd0774be32e7f5d00d9ce8f90ee761b4

Observation c44dbbad-f78c-496b-8e5c-8b778324d395 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:02.518157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:02.518157Z digest=sha256:da38178f9869967162e45630a81ab414a6399f3662c9a953ad80932caa03bffc

Observation 28d60e2c-6b2c-4f67-b468-3f19292ebae9 · outbound

This paper cites Javisdit: Joint audio-video diffusion transformer with hierarchical spatio-temporal prior synchronization.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Javisdit: Joint audio-video diffusion transformer with hierarchical spatio-temporal prior synchronization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:02.653794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:02.653794Z digest=sha256:e8d07a8238181c09fb408ab1012ff123ae4e9255f888dbd3a21f524c767e3829

Observation ba4503df-60bb-47f6-a857-5313a4c1294a · outbound

This paper cites Primitive-based 3d human-object interaction modelling and programming.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Primitive-based 3d human-object interaction modelling and programming

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:12.407256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:02.756572Z digest=sha256:99603f22edcbc88cc28a8ed8542c7f83caa1fe2741eabc29cd3419d42b075826

Observation f0c0aa2d-1aeb-4973-8e75-5fa41bda1d29 · outbound

This paper cites Geneoh diffusion: Towards generalizable hand-object interaction denoising via denoising diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Geneoh diffusion: Towards generalizable hand-object interaction denoising via denoising diffusion

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:12.179360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:02.865643Z digest=sha256:db928dda349554f3d2b2f5d8bc44570551b40b09cd0d86d3d8b5486cf4707c03

Observation 7bc42f5b-ea4c-4053-8ee7-2cdedb949707 · outbound

This paper cites Taco: Benchmarking generalizable bimanual tool-action-object understanding.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Taco: Benchmarking generalizable bimanual tool-action-object understanding

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:02.945310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:02.945310Z digest=sha256:72c7d9e42fd7cfa7a947160f530b6b39f72c8a0d5ac5dd0aefbb36bb77367521

Observation 15c7421d-9dfc-4fcf-ac89-347c671da035 · outbound

This paper cites Hoi4d: A 4d egocentric dataset for category-level human-object interaction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Hoi4d: A 4d egocentric dataset for category-level human-object interaction

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:11.934837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:03.014251Z digest=sha256:987110bc258081dfb49d07acafaad3791f34334de3974482e07de3f1952138b5

Observation 0d7158d8-3810-49d5-b17b-621d14f5f181 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.126140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.126140Z digest=sha256:740b57081e7f0d4da13d1b97ba2f3e0c865fb8f14302ae8332ebaa07a5582d1c

Observation 025c8fac-bb0a-4fe7-a7b7-65b6a93e7be7 · outbound

This paper cites Omnigrasp: Grasping diverse objects with simulated humanoids.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Omnigrasp: Grasping diverse objects with simulated humanoids

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:11.673914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:03.223803Z digest=sha256:eebfedb62994f2fcde0efe4b0145ba03e34c1c8326ed73f5e9060aedc48b2353

Observation f42461c3-a015-461e-af72-4fcd54a1e1c8 · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view synthesis.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Nerf: Representing scenes as neural radiance fields for view synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.376595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.376595Z digest=sha256:b2c1cc9d1c03f0c199e22a3a0dd4db7a915230a4777d4beeff11333a85ecf6ce

Observation b662a0da-3f5e-4073-81af-8693d78100ae · outbound

This paper cites ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.505575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.505575Z digest=sha256:62a776befa34d3c3dd07a24eeb32b3a2bd969b1ae8afce06e9c69ef1c22efd9c

Observation a2198627-a368-4d79-9c2b-c0ef489fbc62 · outbound

This paper cites Scalable diffusion models with transformers.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Scalable diffusion models with transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.652566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.652566Z digest=sha256:47b01e46fd56378bde8c7c70f93589882e80f9540960a1e0258731c7e56de1d5

Observation 93224f88-b4bc-4486-91b4-98431024c3ed · outbound

This paper cites HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.743604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.743604Z digest=sha256:2b2188a06d2d1a6ff10c510ea1435bd954e3436e0b1b297049afecb1663b8636

Observation 1ff6d839-d05f-4233-b385-999e7d9e8697 · outbound

This paper cites Hierarchical generation of human- object interactions with diffusion probabilistic models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Hierarchical generation of human- object interactions with diffusion probabilistic models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:11.441585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:03.835908Z digest=sha256:0504e570f10866aaa8cf18dd9e656f8b9e7d30b9d2523b35f322f98b5cb3f1d7

Observation 6b567976-c336-4f54-aa2c-799edab2aa0c · outbound

This paper cites Learning transferable visual models from natural language supervision.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Learning transferable visual models from natural language supervision

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.951478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.951478Z digest=sha256:f15fb4a6e9194912e4adf2e7fe7d43b5c1a5aadb2573155689f2f3091c47f7a5

Observation 2bb6fe21-95fe-478b-a244-617248d39e83 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.037964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.037964Z digest=sha256:e7ccd96a9f7bd9b88aeac2e60830c744bd78b52a9d1cf8fff6068fd9141dc435

Observation 38c69bab-f3bb-4ad7-872f-d312d59ab737 · outbound

This paper cites Zero: Memory optimizations toward training trillion parameter models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Zero: Memory optimizations toward training trillion parameter models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.208113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.208113Z digest=sha256:35a014f5a9708b0b0666d78479fa190834abaa29b5b8049e13388468984547e6

Observation 270da79e-0c15-4a3c-96af-7fa997ca26ee · outbound

This paper cites Zero-shot text-to-image generation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Zero-shot text-to-image generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.351168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.351168Z digest=sha256:f5722ac2b7ef9a1ca4a6da89e1273e279e5d7c2567e252dabe0bfccd8d856393

Observation 5499978a-9eff-4e8f-9a75-3232ac3ca477 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios High-resolution image synthesis with latent diffusion models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.461811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.461811Z digest=sha256:905fe0c8554c7800c73df929b651706e94d523e7dfe20fcfe5cf1c003a3646c0

Observation 600a4446-509d-4136-b0ea-d0fc9cdc6c4d · outbound

This paper cites Photorealistic text-to- image diffusion models with deep language understanding.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Photorealistic text-to- image diffusion models with deep language understanding

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.570786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.570786Z digest=sha256:f5c56327007ec993edd1b09c87e74b86baa5e669fb29b82cf395140a41bcc6ac

Observation 5ed1133c-a3bf-4f78-b182-ae50846e14a6 · outbound

This paper cites Hand-Object Interaction Pretraining from Videos.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Hand-Object Interaction Pretraining from Videos

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.683935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.683935Z digest=sha256:14f81d9e2f9930a78892253ff09c5a731ff121274fd14e36a3fa13ebc3d4f57e

Observation 50def3c1-627d-4f21-9cba-998e4202f6ef · outbound

This paper cites Grab: A dataset of whole-body human grasping of objects.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Grab: A dataset of whole-body human grasping of objects

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:11.084079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:04.800452Z digest=sha256:88f94013dc67db560918d50333cf2d7114f073cff2b5f7fc1825a1b7348a3caa

Observation c3803749-a5dd-449b-bc5f-ada3713a056b · outbound

This paper cites Any-to-any generation via composable diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Any-to-any generation via composable diffusion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:10.833308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:04.885671Z digest=sha256:57e1c85a65b4b576d37bc2c86d0dd36d99dd2ac0c4056e567a337c02acc9ab4f

Observation b29a7c5c-0251-4a74-b1b3-e00ba1d0f11e · outbound

This paper cites Raft: Recurrent all-pairs field transforms for optical flow.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Raft: Recurrent all-pairs field transforms for optical flow

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:04.953168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:04.953168Z digest=sha256:466c1ae95fcd6f47041f9fe035ba629791ed9f457c73eff575f04357d8cbed6e

Observation c2afb283-976e-4948-bcd5-32f5e6efde19 · outbound

This paper cites Human motion diffusion model.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Human motion diffusion model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.028705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.028705Z digest=sha256:be7e0ffaa389c21a71fa300056e9925509f3506809a9c7a03ea7faa34bf446e7

Observation 08143ec8-ca17-4dfb-8927-5e4083f58fb1 · outbound

This paper cites Deepsimho: Stable pose estimation for hand-object interaction via physics simulation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Deepsimho: Stable pose estimation for hand-object interaction via physics simulation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:10.543054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:05.104217Z digest=sha256:e68b46c83ee960efcb9a3c1a77d0e55d16904aecc483e884c833d4f82c6f05e5

Observation d19071c2-fc6a-4069-b907-8c35f059d613 · outbound

This paper cites Cogvlm: Visual expert for pretrained language models.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Cogvlm: Visual expert for pretrained language models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.143860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.143860Z digest=sha256:5b10ed58cdd08baa94db48eccba52f75665e48189618d1c03cb0af8e547c540e

Observation 7d32ac8f-6ed1-4526-9d2e-d7aa85ee4bad · outbound

This paper cites PhysHOI: Physics-Based Imitation of Dynamic Human-Object Interaction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios PhysHOI: Physics-Based Imitation of Dynamic Human-Object Interaction

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.220366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.220366Z digest=sha256:2c2458606ed7cc83afd7a7c05477cbdbf6aef779efc0222272f9b6e4fefe5e91

Observation 50bdc760-77cb-44a8-8dbf-ef1cc691b94a · outbound

This paper cites Easyanimate: A high-performance long video generation method based on transformer architecture.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Easyanimate: A high-performance long video generation method based on transformer architecture

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.314134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.314134Z digest=sha256:b61c405fb569af4a7d4b9949fc4463982115559a7fe9c5802d4a7aa6c5e2b267

Observation 5a401baa-84e2-423e-9cf6-f8262337e588 · outbound

This paper cites Interdiff: Generating 3d human-object interactions with physics-informed diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Interdiff: Generating 3d human-object interactions with physics-informed diffusion

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.380924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.380924Z digest=sha256:3cb7c243fff52aee973c7ddec29ea4fa63b0aa9ebb244e0fad4ff640371d2e39

Observation e62e22f1-2c2e-47be-99ec-8931538b9667 · outbound

This paper cites Intermimic: Towards universal whole-body control for physics-based human-object interactions.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Intermimic: Towards universal whole-body control for physics-based human-object interactions

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:10.325426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:05.471978Z digest=sha256:f03013eeaed83f903556cc23f4b39165e76ec0600ddc42f1620abee668a49d3d

Observation def8529f-e32e-4ab7-8cb7-e31f92890363 · outbound

This paper cites Interdreamer: Zero-shot text to 3d dynamic human-object interaction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Interdreamer: Zero-shot text to 3d dynamic human-object interaction

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:10.117411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:05.546487Z digest=sha256:4388a12c9d44cf59eba4ea265e2938633d49174dc481e3686de6205792b56dc4

Observation 772d192a-2603-4302-838c-3c549f6d46c5 · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Magicanimate: Temporally consistent human image animation using diffusion model

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.632537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.632537Z digest=sha256:34d6890d6ee3e18bca13a625833b6fac2dbfff344b607b96b52cf72149a6375f

Observation 96c659a6-e3e8-4831-98cc-1275780fb212 · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.702331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.702331Z digest=sha256:f0a814beec4171f3dd0ddeeb5261bb7da7849a213a7f61f5f89dc322928f084b

Observation 84744637-242e-4531-bfcb-479fa12ea8cf · outbound

This paper cites Oakink: A large- scale knowledge repository for understanding hand-object interaction.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Oakink: A large- scale knowledge repository for understanding hand-object interaction

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:09.810360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:05.797285Z digest=sha256:8cb497bea2843332cdb167e93cf860c244055a90b1396dd4c6b15e944bf25e99

Observation fc7a88ea-b638-4a6f-ab1d-2ad03c9da93a · outbound

This paper cites The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision).

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:05.840015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:05.840015Z digest=sha256:c04ed5db03b5f3a7f5fe7798eb7d3faa0d0b542ebfd8ba03e0fbc7eb33a290e9

Observation 3a7b274a-0c18-4a97-b69e-09c546b768a9 · outbound

This paper cites Cogvideox: Text-to-video diffusion models with an expert transformer.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Cogvideox: Text-to-video diffusion models with an expert transformer

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:09.542934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:05.870422Z digest=sha256:b69812dae4937641af781d1a3326373da8c2a27a8f676042d38d1b7e47dd4a0d

Observation 864b99ea-f983-4e63-bbd3-0784000f2e8a · outbound

This paper cites Diverse and aligned audio-to- video generation via text-to-video model adaptation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Diverse and aligned audio-to- video generation via text-to-video model adaptation

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:09.302029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.010511Z digest=sha256:71d38d16ad15d618d6563f03f807709bfb7a3063597fcddfb910ca3dafc91643

Observation f2d41a80-24d2-4dac-a3ff-2acb967e47d9 · outbound

This paper cites Oakink2: A dataset of bimanual hands-object manipulation in complex task completion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Oakink2: A dataset of bimanual hands-object manipulation in complex task completion

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:08.986003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.121308Z digest=sha256:a42ce80e136089f5f631ca20d15685bef7c13897d4b875bfd6a6dc8f62047c6a

Observation 34b26231-8c05-45e5-9c00-b99e66179bbb · outbound

This paper cites ManiDext: Hand-Object Manipulation Synthesis via Continuous Correspondence Embeddings and Residual-Guided Diffusion.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios ManiDext: Hand-Object Manipulation Synthesis via Continuous Correspondence Embeddings and Residual-Guided Diffusion

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:29:07.077298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.314273Z digest=sha256:6de44e3cbc70def0e9123e35323f94d13717b5858ca531d3b87eb8c666d26586

Observation f2db1453-ab85-47f9-97a0-0bb598b0199e · outbound

This paper cites Couch: Towards controllable human-chair interactions.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Couch: Towards controllable human-chair interactions

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:08.711736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.492979Z digest=sha256:20af6182a629bde915a58bfe6b538cae947e1b921d4066f907034b7a0377a8e3

Observation 8ff90525-bead-41b8-b02f-8533fb05c192 · outbound

This paper cites Emdm: Efficient motion diffusion model for fast and high-quality motion generation.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Emdm: Efficient motion diffusion model for fast and high-quality motion generation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:08.378340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.643675Z digest=sha256:c1abc87a9e8b87a976ecf8ecc0b23f2c381228c243ec1064ef30d8844d0624ba

Observation 3b4a5d85-8cb1-42a5-98ed-9a6c23b3703f · outbound

This paper cites Champ: Controllable and consistent human image animation with 3d parametric guidance.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Champ: Controllable and consistent human image animation with 3d parametric guidance

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:08.084227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.800490Z digest=sha256:9ad3c5684ad299a8a77e711b626a21a0d61440af165baaf0f45022c15a9a2064

Observation 06640dcd-bcb3-4c4b-92ab-e3a65c68a1ad · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios Rt-2: Vision-language-action models transfer web knowledge to robotic control

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:07.738463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:29:06.878705Z digest=sha256:bed274a058ea06d11212ddc010aced23e09d710ff94e27bf87e1b91f9e6ef0b9

Pith citing papers

Observation 3e3cca65-46dd-480f-ade2-e5256b5c6423 · inbound

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis cites this paper.

HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:18:30.110150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T00:14:36.537688Z digest=sha256:43ddea1d95383a9935329c993e45f33aa2ce479863e6a771519e2711ec4f41fb

Observation f8beea9c-8d04-4da1-b448-79dfe4f7dd5b · inbound

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation cites this paper.

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T10:37:55.063859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:37:55.063859Z digest=sha256:a5fbe650c916a76d86791098892775654c502ced0a3d1dff1bbacb9315d52a3f

Observation a81105ac-ac46-4737-99dc-576b15490232 · inbound

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment cites this paper.

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T05:27:16.175479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:27:16.175479Z digest=sha256:b140789192317258136b5671d146d928da569d4f8f159794311073ea90affa10

Observation c76c5d80-b73d-4dfc-8da3-743d5c59f131 · inbound

PhotoHOI: Synthesizing 3D Hand-Object Interactions from a Single RGB Photograph cites this paper.

PhotoHOI: Synthesizing 3D Hand-Object Interactions from a Single RGB Photograph SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T18:35:26.468228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:35:26.468228Z digest=sha256:f21f100dfabad1301a275b11d3703d3cb4348a325176838b5a4816c6ddfa104e

Observation 0cd1448c-0687-4bd5-9198-157da069b78e · inbound

PhotoHOI: Synthesizing 3D Hand-Object Interactions from a Single RGB Photograph cites this paper.

PhotoHOI: Synthesizing 3D Hand-Object Interactions from a Single RGB Photograph SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:14.672688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:10:14.672688Z digest=sha256:811aab501948388bff476c4657a0183ab02fc1edf05e4ec666f9606c821393fa