Pith. sign in

Paper Citation Record · LEDGER

HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2312.06553.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.06553 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:32:42.870119Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T16:17:21.682636Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5193cf69-ff94-4e03-86f9-1980e818765c · inbound

Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression cites this paper.

Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-12T13:02:31.932191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:02:31.932191Z digest=sha256:7335182b2bc8d29fe464845f6915c563e3f46c1cde9869a769f9708c7a072faf

Observation f52f068a-ab2f-47e5-b585-9cc01853894c · inbound

OOD-HOI: Text-Driven 3D Whole-Body Human-Object Interactions Generation Beyond Training Domains cites this paper.

OOD-HOI: Text-Driven 3D Whole-Body Human-Object Interactions Generation Beyond Training Domains HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T11:29:41.090807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:29:41.090807Z digest=sha256:32946e20a35eb830878ea22b8e80a55655337cd5691c6f2976e8749f2c32e577

Observation e5fe6326-7b0e-423a-b5cf-0ea0d3fcdf3f · inbound

BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects cites this paper.

BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T21:05:12.277085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:05:12.277085Z digest=sha256:6d3400507f2a9c409087042b1747e0d72822e635a58de1e3afb8583fc4ff77ad

Observation bfd11821-77d1-4768-ba44-d9af421ad446 · inbound

MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow cites this paper.

MulSMo: Multimodal Stylized Motion Generation by Bidirectional Control Flow HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:40.686409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:40.686409Z digest=sha256:3d226b5e55ab81171e5d79b2d1f270e471cf0e4525492d9c60829ef3344de6e8

Observation ea7bc102-55ac-49ee-bf5c-1680a04b0d0d · inbound

SCENIC: Scene-aware Semantic Navigation with Instruction-guided Control cites this paper.

SCENIC: Scene-aware Semantic Navigation with Instruction-guided Control HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T11:16:28.264747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:16:28.264747Z digest=sha256:3ed3f693d9a3892de7d0d3649e599ce44801249007d1309e6264307a73f3bef5

Observation 9c231415-547c-4fbc-a454-b27c14b07eb0 · inbound

Mimicking-Bench: A Benchmark for Generalizable Humanoid-Scene Interaction Learning via Human Mimicking cites this paper.

Mimicking-Bench: A Benchmark for Generalizable Humanoid-Scene Interaction Learning via Human Mimicking HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T05:19:19.027080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:19:19.027080Z digest=sha256:ce8f7096268eccdd6ef78ae552b5b7a19ab9edb88f63a912fb42bbbb9b4b507c

Observation 8eb43e31-caf9-442b-97fc-7a5d2aba0b10 · inbound

SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis cites this paper.

SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T23:38:23.764585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:38:23.764585Z digest=sha256:3931a2244be33d08f387d660e38c3aaaedc60eb8af1c7783a4f8cde466b95c3b

Observation 027ea1e1-5a3f-47c8-a997-c7b9f29e10a4 · inbound

Diffgrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion Model cites this paper.

Diffgrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion Model HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T23:21:18.406269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:21:18.406269Z digest=sha256:42052be5e78451c1c106e9cdfd97841d608affc1f8ba58199d1226f57960c80f

Observation c6623e27-b173-40a2-b679-105538153574 · inbound

UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes cites this paper.

UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:32:42.870119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:32:42.870119Z digest=sha256:3d85f173e2d2f15a383bec37be9b72bc42cef0b0df8f96c8390b95d3e25cbe81

Observation b2beab20-8572-43d2-8d0c-52e003fde50a · inbound

Absolute Coordinates Make Motion Generation Easy cites this paper.

Absolute Coordinates Make Motion Generation Easy HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:18.087763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:19:18.087763Z digest=sha256:1b464ebe9bab511710ad51d560de6aa72bab66a55237012a8f4ac0820e9cea99

Observation afe65e7d-b9db-475c-ac71-48bb6fa55ac3 · inbound

CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects cites this paper.

CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:47.245280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:47.245280Z digest=sha256:a13652b21568f823c13e0b2c3c9930760b7aeb8d5518a7f7dee3d8c83d92d75b

Observation 0046c9ea-27db-42b8-b70d-bddd04d9cdc0 · inbound

InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing cites this paper.

InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:26.726400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:34:26.726400Z digest=sha256:cf11feec05564bafb57540bce385a89a19aa29da4b462d7d537f0962dfdf6bb8

Observation 93224f88-b4bc-4486-91b4-98431024c3ed · inbound

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios cites this paper.

SViMo: Synchronized Diffusion for Video and Motion Generation in Hand-object Interaction Scenarios HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:03.743604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:03.743604Z digest=sha256:2b2188a06d2d1a6ff10c510ea1435bd954e3436e0b1b297049afecb1663b8636

Observation 1426234f-cc75-487c-bd24-345c3c0d8d81 · inbound

InterMamba: Efficient Human-Human Interaction Generation with Adaptive Spatio-Temporal Mamba cites this paper.

InterMamba: Efficient Human-Human Interaction Generation with Adaptive Spatio-Temporal Mamba HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:53.069776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:53.069776Z digest=sha256:dc0cd18bb025500c5d9052caa6997d0ec4eba0cd0ceb900ae52995f1c71b6115

Observation 4dcea8d2-cb25-4ec2-9951-010cb855c598 · inbound

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation cites this paper.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.105677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.105677Z digest=sha256:56f42830880ecc2a55262a58b1995b688bae85a5fa08209b94652e613661ea66

Observation aa288ba1-566b-4548-b58f-0968803808ee · inbound

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers cites this paper.

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:39.652056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:27:39.652056Z digest=sha256:27abc4f9080335be7a3c5c6a1ab82568b0715f524e8cc3e4d4747ccfca813cd3

Observation 31686fc4-93d8-4b42-87e7-0e45c9473103 · inbound

GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects cites this paper.

GenHOI: Generalizing Text-driven 4D Human-Object Interaction Synthesis for Unseen Objects HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:45.792596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:58:45.792596Z digest=sha256:910b89daf37f6afbf84d4e01466f21c39af00ea7c13b4b65da254b2c646cbde8

Observation 5591942f-fc0b-4541-b148-251f5b228a7d · inbound

GenHSI: Controllable Generation of Human-Scene Interaction Videos cites this paper.

GenHSI: Controllable Generation of Human-Scene Interaction Videos HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:27:09.035768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T07:25:41.219753Z digest=sha256:60a1211440e35b259e4830630c1a72ca645949fec90a81c5c2c741f220bcfaad

Observation 8d0e308e-1883-4061-a35a-d07bf788daf1 · inbound

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions cites this paper.

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T23:53:46.456426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:53:46.456426Z digest=sha256:24b1fba9a1a9f99a00a29c2d99d8f3c6f3322048bd678fb4f1b9ffdbbd19658b

Observation 26a67449-2df7-4bab-8dfa-32246db3469a · inbound

SimGenHOI: Physically Realistic Whole-Body Humanoid-Object Interaction via Generative Modeling and Reinforcement Learning cites this paper.

SimGenHOI: Physically Realistic Whole-Body Humanoid-Object Interaction via Generative Modeling and Reinforcement Learning HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T17:21:55.409054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:21:55.409054Z digest=sha256:fffd8bd25bff4697f150234a0d4509a4c123eca742c6517f1d8e0bac2bf512f1

Observation 7e21e662-905d-4a66-8073-cebb53688121 · inbound

ECHO: Ego-Centric modeling of Human-Object interactions cites this paper.

ECHO: Ego-Centric modeling of Human-Object interactions HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T16:44:38.690640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:44:38.690640Z digest=sha256:64f256d7989f10900f568785b90441ed55b22d67c4d0c1fbd56352b887361c5a

Observation 4a82ca20-6636-4d4f-989d-db8bf0622d07 · inbound

ScoreHOI: Physically Plausible Reconstruction of Human-Object Interaction via Score-Guided Diffusion cites this paper.

ScoreHOI: Physically Plausible Reconstruction of Human-Object Interaction via Score-Guided Diffusion HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T16:14:01.877696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:14:01.877696Z digest=sha256:d17267d164570e59411eb831b0046582d0a193b158e1cad79bc1eda0de32e40d

Observation 5b7ad611-f3a8-4f53-98bb-90305b69002f · inbound

InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation cites this paper.

InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T18:57:10.451963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:57:10.451963Z digest=sha256:5281c456e96923b9b4bd75107f47e814f19210e5f5a75727082dec7c19b6517f

Observation 87a17c1a-b204-415f-bfc0-3f0bd541b603 · inbound

Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation cites this paper.

Coordinating Multiple Conditions for Trajectory-Controlled Human Motion Generation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:57:52.888299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T19:57:45.837431Z digest=sha256:3c42f004d0c2330b2b56d2e29be11645346209453c84812554fc7bea401c69e6

Observation 97a7cc1a-7cbe-40d7-8ecf-6b0ab8afd553 · inbound

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics cites this paper.

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:39:46.890362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T09:40:37.183422Z digest=sha256:a59ec95bc5d4bf3c386f0d2e71653da99056b1972e17705d2032792e046325e4

Observation ecb4114b-d8fd-432b-9de4-5def0437ed19 · inbound

GIRAF: Towards Generalizable Human Interactions with Articulated Objects cites this paper.

GIRAF: Towards Generalizable Human Interactions with Articulated Objects HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-10T16:17:21.683944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-10T16:07:56.971414Z digest=sha256:ba8a0dfd3a4e4de463c31992245342bfdb225ab997a1072827fb75c7f0f8e0c4

Observation a49675f6-cc17-413b-aabc-61fcfe460b72 · inbound

ContactMimic: Humanoid Object Interaction via Contact Control cites this paper.

ContactMimic: Humanoid Object Interaction via Contact Control HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-10T02:16:42.501543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-10T02:11:02.007137Z digest=sha256:62ad1d4a350b2d80900f619088f20d87843fbf6fb12f44a1cdcd4e1fa100b98a

Observation 2198be78-3636-442e-aae1-c9dfa8946ba4 · inbound

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis cites this paper.

HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 164

Resolution
unresolved
no resolver link, observed 2026-08-01T19:06:44.892727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T19:06:44.892727Z digest=sha256:35d26ed6a11f4b217ca6c9a70b2777bc9d5166747b1482e202b14483840cc5ba

Observation 40667efe-ba8b-434e-91b5-098c1ededd18 · inbound

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment cites this paper.

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T05:27:18.517879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:27:18.517879Z digest=sha256:76679cd9e4ac5729198980cddfbe1337ba1703cbb2cee4873a231fac1e5a847a

Observation 9705e7ae-ef88-4b4d-a51e-e406836eec26 · inbound

MoRAE: Flow-Friendly Self-Supervised Latents for Text-to-Motion Generation cites this paper.

MoRAE: Flow-Friendly Self-Supervised Latents for Text-to-Motion Generation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T12:14:43.657566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:14:43.657566Z digest=sha256:acbef42672e22b5c3f4f0e420004cc3280449f05d2fb0cba6a2a14307bd59128

Observation b9e1abeb-7bd4-4574-99a0-10f399cddd7b · inbound

Surface Keypoint Representation for Multi-Object and Articulated Human-Object Interaction Generation cites this paper.

Surface Keypoint Representation for Multi-Object and Articulated Human-Object Interaction Generation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-15T14:58:47.344161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:58:47.344161Z digest=sha256:88d523b9542438c99f8a84b091e859d5142f160726739a498c9e4289e0a13017