Pith. sign in

Paper Citation Record · LEDGER

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2505.06832.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06832 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:35:20.174431Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation abffebaa-f554-412f-ba87-205573280994 · outbound

This paper cites Open-vocabulary queryable scene representations for real world planning,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Open-vocabulary queryable scene representations for real world planning,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.063393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.063393Z digest=sha256:03143f5eff46728d6b09e2e20e169e9b0f647213de5d43449ae9e7e936ffd207

Observation 56dc0305-d652-4ed7-9565-e285df3893c3 · outbound

This paper cites A joint modeling of vision-language-action for target- oriented grasping in clutter,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms A joint modeling of vision-language-action for target- oriented grasping in clutter,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.068085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.068085Z digest=sha256:acc93871f0be4d69f82eaf2999fc8e21bb4e152597220407c2832f266c0ab71b

Observation bfe5c818-2297-4709-89ea-f7732c4a28b8 · outbound

This paper cites Da 2 dataset: Toward dexterity-aware dual- arm grasping,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Da 2 dataset: Toward dexterity-aware dual- arm grasping,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.512527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.072033Z digest=sha256:674951d5b023e7ffaf3edbb93784e7bd600a0bfc0d8ee29a3b8f7212464cf29a

Observation a2db953e-e181-45b5-b288-889b134914a9 · outbound

This paper cites 6-dof graspnet: Variational grasp generation for object manipulation,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms 6-dof graspnet: Variational grasp generation for object manipulation,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.076427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.076427Z digest=sha256:52493810893d822d9851dcb649d63ad5b3fb2e59ad1f79322fa12a979bb46b0d

Observation c11e139c-64e2-46b5-9bd9-24dc5c5629f7 · outbound

This paper cites Contact- graspnet: Efficient 6-dof grasp generation in cluttered scenes,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Contact- graspnet: Efficient 6-dof grasp generation in cluttered scenes,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.080418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.080418Z digest=sha256:d77f5385e88c5f4ae01c6b3edd88d3d04baf6a2d8cc3cd345d5be9fb30693341

Observation 3074b412-0c27-4484-9c9a-c59427da2d52 · outbound

This paper cites AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.084292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.084292Z digest=sha256:5e436e5a48be1cc87794eede642cd22a34166b9fbd827b780017cbe333e38f52

Observation 52e66a74-a920-4f27-b1d4-1c902c17d0a3 · outbound

This paper cites Thinkgrasp: A vision-language system for strategic part grasping in clutter,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Thinkgrasp: A vision-language system for strategic part grasping in clutter,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.088798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.088798Z digest=sha256:ada62eb443d7d62678032ef8b211080e26c92eab489a1aeac552a3e7244f8e58

Observation 1fc182d7-6f74-4c5b-b299-f8ce5bf47f5b · outbound

This paper cites Graspnet-1billion: A large- scale benchmark for general object grasping,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Graspnet-1billion: A large- scale benchmark for general object grasping,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.092396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.092396Z digest=sha256:bb9861ac6db657c8febd72d3e4dc3de5aa7c9a3c67349e526303a84e3bce3e80

Observation 7eadc6af-b5e5-4a31-86f6-94a68f7897bc · outbound

This paper cites Anygrasp: Robust and efficient grasp perception in spatial and temporal domains,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Anygrasp: Robust and efficient grasp perception in spatial and temporal domains,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.096056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.096056Z digest=sha256:76f0bdf8c3a897247911f02129dc4330769bfd6f83c276202810596c7f0355f0

Observation 8714024c-41b6-4407-a612-c133b6089ee6 · outbound

This paper cites Constrained generative sampling of 6-dof grasps,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Constrained generative sampling of 6-dof grasps,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.470871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.100086Z digest=sha256:0459b4c387da3e9b9b95ab297b7bf8595ff98e46b3e6159c2710f355fadf43f7

Observation 91e3b5c5-340c-4ca5-8957-3ab2282a498e · outbound

This paper cites Constrained 6-dof grasp generation on complex shapes for improved dual-arm manipulation,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Constrained 6-dof grasp generation on complex shapes for improved dual-arm manipulation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.459107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.103765Z digest=sha256:6664ba36c4c59b258593523125d139f52a4e44ecce46a1a36408989d5f22703a

Observation bcbe0fa7-ac19-4ffb-b063-2c7920809d87 · outbound

This paper cites An overview of 3d object grasp synthesis algorithms,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms An overview of 3d object grasp synthesis algorithms,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.447154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.107464Z digest=sha256:6784a419578c137856e2b43b57619a8f817d4f45132b3947600b62da487d3026

Observation 697d5490-5742-45d6-97b2-8dec02e0ddc2 · outbound

This paper cites Affordances from human videos as a versatile representation for robotics,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Affordances from human videos as a versatile representation for robotics,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.111778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.111778Z digest=sha256:7f92964561fdf710dbdfac45e170600eb9bae7dbbecd55569aa9f0b87b16f3f7

Observation 25765b6d-f877-4580-99f4-1c2ad1e90443 · outbound

This paper cites Robo-abc: Affordance generalization beyond categories via semantic correspon- dence for robot manipulation,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Robo-abc: Affordance generalization beyond categories via semantic correspon- dence for robot manipulation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.115971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.115971Z digest=sha256:362c40bb73827f78c275e70362decee2ac90eb63f6787fc3664558e01e6c843c

Observation 58dcf9bf-be08-4cb6-9b25-0fe4a19aaedf · outbound

This paper cites Graspgpt: Leveraging semantic knowledge from a large language model for task- oriented grasping,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Graspgpt: Leveraging semantic knowledge from a large language model for task- oriented grasping,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.120173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.120173Z digest=sha256:5c628cf2c97a65b187e359adac74782e85365537f28aa10a4b47e658dc432e24

Observation fe5579c2-4414-464c-84dd-892de7ab9953 · outbound

This paper cites Foundationgrasp: Generalizable task-oriented grasping with foundation models,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Foundationgrasp: Generalizable task-oriented grasping with foundation models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.413631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.124051Z digest=sha256:be2c87beb33bbed8a8ad5542ba54e08cdacd4046b2b7c159de3ae450f415a84a

Observation 858d775f-05ce-4ec9-bb41-abe6f6fc6d01 · outbound

This paper cites Affordance grounding from demonstration video to target image,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Affordance grounding from demonstration video to target image,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.399015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.127931Z digest=sha256:d9c5188836195f918fe304081f9bfe0b28787a8f2607d6699cd3419aed0c3456

Observation 71a85292-8b2c-404b-9d79-ff193c71a4c2 · outbound

This paper cites GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.131854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.131854Z digest=sha256:ac79e370ca0cb7a7df23022b454c76e9c2bd6763d0b3f433751fa790b72b2cdc

Observation 9f3c6c09-6eb5-4b17-b174-c982910a4761 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Grounding dino: Marrying dino with grounded pre-training for open-set object detection,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.385613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.140240Z digest=sha256:befb239771961d45ab57cbb758f5a6a03a449041e527fbf52defb7ef2831d54d

Observation 844ffc8d-24b3-46b8-8f30-22458cf4d6b5 · outbound

This paper cites Segment anything,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Segment anything,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.373788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.144324Z digest=sha256:b4d6a3ae1b2de82ffbd3312de274a8fb1d907bdb0782b9fa4404bd6a79d8998a

Observation 44480771-5276-4b38-95a4-2aa7f3df8aed · outbound

This paper cites Going denser with open-vocabulary part segmentation,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Going denser with open-vocabulary part segmentation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.362382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.147997Z digest=sha256:4eeeadcaf3e6a20e5b0cd611fd14c4451de77735ae739475e265972a6cc45b27

Observation a704b519-1b77-4792-9a5a-7fc341c52c8f · outbound

This paper cites Lan- grasp: An effective approach to semantic object grasping using large language models,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Lan- grasp: An effective approach to semantic object grasping using large language models,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:20.350584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T22:35:20.151482Z digest=sha256:6cd2994a74a1e8197dfd746266d64a57856a909578cce42f65304de5f3e13b4f

Observation c5667adb-9059-4247-ab98-b2d42df0c5d3 · outbound

This paper cites OVAL-Prompt: Open-Vocabulary Affordance Localization for Robot Manipulation through LLM Affordance-Grounding.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms OVAL-Prompt: Open-Vocabulary Affordance Localization for Robot Manipulation through LLM Affordance-Grounding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.155019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.155019Z digest=sha256:2cc151bfd4321ff0bb9cb1275d99d801b668142d38c99ed10d4ce16a722ae00f

Observation fe9bea7f-70ae-4f52-9b2d-92492088e56d · outbound

This paper cites Se (3)- diffusionfields: Learning smooth cost functions for joint grasp and motion optimization through diffusion,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Se (3)- diffusionfields: Learning smooth cost functions for joint grasp and motion optimization through diffusion,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.158622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.158622Z digest=sha256:698bfaabcc18e39e4cbe7dbcecca0097f360afc36148069b6e89dc71c86ea92e

Observation 2cfaa855-9c87-4996-9b66-5bce6ac59f36 · outbound

This paper cites Convolutional occupancy networks,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Convolutional occupancy networks,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.162414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.162414Z digest=sha256:b04d067333b2760f6834d597982747b87b0c8b1a376a95908f905b70acd4b6d0

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.166200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.166200Z digest=sha256:1457c3584192ec5dfe6597d1d52e5cb81d3ead62d8671f53694eca0d833e2ab5

Observation 0cc98b6a-7029-4e72-a173-9bf34f8cd566 · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.170182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.170182Z digest=sha256:34ea7eca0775f1f3b53ad4c6ded8a20cd91036278c6539248ad3a77e33fdf133

Observation a08776a3-2dc9-4e12-a5c0-0ddeb692a585 · outbound

This paper cites Synthesizing diverse and physically stable grasps with arbitrary hand structures using differentiable force closure estimator,.

UniDiffGrasp: A Unified Framework Integrating VLM Reasoning and VLM-Guided Part Diffusion for Open-Vocabulary Constrained Grasping with Dual Arms Synthesizing diverse and physically stable grasps with arbitrary hand structures using differentiable force closure estimator,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:20.174431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:20.174431Z digest=sha256:c1799405163f6dc099647a0373f9752215aac55fc75b7b11547e72c651bd61a4

Pith citing papers

No inbound Pith citation observations are available.