Pith. sign in

Paper Citation Record · LEDGER

DexVLG: Dexterous Vision-Language-Grasp Model at Scale

As of 14 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 8 inbound Pith citation observations for arXiv:2507.02747.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02747 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:29:33.097470Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:33:39.445246Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T06:04:34.676226Z

Reference resolution

73 of 73 outbound references displayed

  • verified exact3
  • verified fuzzy35
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 10f0d90b-4765-4165-b495-0467919d6626 · outbound

This paper cites GPT-4 Technical Report.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:23.959143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:23.959143Z digest=sha256:94bdbc0af8e57c4cec986e0d8f7ed681566830898beea97bd1f8700b95f7e781

Observation 0c4f6461-ba29-48e5-9111-f06bfd3539c9 · outbound

This paper cites Dexterous functional grasping.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexterous functional grasping

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.919023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:24.089645Z digest=sha256:fcd18957cc11333cd5bbe8d0ac72be88b52e439000831f8c8a87e65d22d89e5d

Observation ede89f73-b332-4bb4-b453-7dec4abab9d5 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.283858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.283858Z digest=sha256:0a3be5c2cfc1f61366b85723fd40464f4931d1dc5298181a0ab8d2f61cf7fc14

Observation dfcc98e9-132b-4609-a7d5-71429ee5c4e2 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.467366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.467366Z digest=sha256:aa90a2feabf702e0effec5678988a8888813405ceec7bbedfb6a4377c6ff26ed

Observation a15c7294-a4df-49b9-ac51-e115657935ca · outbound

This paper cites Task-Oriented Dexterous Hand Pose Synthesis Using Differentiable Grasp Wrench Boundary Estimator.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Task-Oriented Dexterous Hand Pose Synthesis Using Differentiable Grasp Wrench Boundary Estimator

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.623262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.623262Z digest=sha256:2393730e080cf2769363b22d8d62d02bde23f706e7fc2c53074d1e9b1d1bc6b1

Observation 60e7b811-93d2-43be-ad01-0fa11a225596 · outbound

This paper cites BODex: Scalable and Efficient Robotic Dexterous Grasp Synthesis Using Bilevel Optimization.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale BODex: Scalable and Efficient Robotic Dexterous Grasp Synthesis Using Bilevel Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.749825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.749825Z digest=sha256:74ef81400ee4b0c366db53fbcc197e498966a04b9803f8267b2442e2dcccabc3

Observation 83ecaee0-2962-41ff-8a05-b3960414cf34 · outbound

This paper cites SpringGrasp: Synthesizing Compliant, Dexterous Grasps under Shape Uncertainty.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale SpringGrasp: Synthesizing Compliant, Dexterous Grasps under Shape Uncertainty

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.877321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.877321Z digest=sha256:b96e41e56270a5024833e8a6575e77539136e625beac72856696b17a90f04007

Observation 60e8948a-ac7d-47c4-b74d-5e1f47b1063a · outbound

This paper cites Learning Robust Real-World Dexterous Grasping Policies via Implicit Shape Augmentation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Learning Robust Real-World Dexterous Grasping Policies via Implicit Shape Augmentation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:25.009478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:25.009478Z digest=sha256:a77318c93731978c48a87f626ce50ae7aa283dadbe070c2e444da21c60274663

Observation 11dfaaf7-284c-492b-9360-c96c594889e4 · outbound

This paper cites Syn- thesis and optimization of force closure grasps via sequential semidefinite programming.Robotics Research: Volume 1, pages 285–305, 2018.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Syn- thesis and optimization of force closure grasps via sequential semidefinite programming.Robotics Research: Volume 1, pages 285–305, 2018

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.786579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.178223Z digest=sha256:d78fd6cac5bafeb0eeab98d7143d4a2b9451ab91760fedc343008080328c01d4

Observation c8bb98b4-d7fe-44a9-9a6c-e77b02dd11b5 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:41.670671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.316048Z digest=sha256:8bb6efdc1223aa9198cb6a5c5b1965ebafcf098d254fe185cb957235f7702ccc

Observation be662dab-1b63-43c1-98ea-fff45681b87e · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Objaverse: A universe of annotated 3d objects

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.529274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.408782Z digest=sha256:971a0d9ece5ea0332dc82e44e34b07e10a2f02fbf393c063b624738a7af96de7

Observation 11c6b362-0cd5-40d8-a3af-8de5df796ed0 · outbound

This paper cites Open6dor: Benchmarking open-instruction 6-dof object rearrangement and a vlm-based approach.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Open6dor: Benchmarking open-instruction 6-dof object rearrangement and a vlm-based approach

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.246754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.517974Z digest=sha256:1e456c64668428e2cd5843a2795a569e9b2718b3524a5cee02a2b388d7325eea

Observation ee315b7d-9b6b-46ff-a9b3-46f4cd2896b2 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:40.986668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.587196Z digest=sha256:f05303cf7a11a8b8659a19c2454ab3371d8b3c98fa82a515484f00779d8e4385

Observation 5a7b652c-04d2-4b32-abec-e36d42a1700b · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale An image is worth 16x16 words: Transformers for image recognition at scale

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.865746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.708066Z digest=sha256:8d6c3c09d7b535994fe361a0670348a89fd71e949653e3c111862b10d953ad3f

Observation 9129cffc-20ea-4882-8833-ecb64fac30c9 · outbound

This paper cites Graspnet-1billion: A large-scale benchmark for general ob- ject grasping.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Graspnet-1billion: A large-scale benchmark for general ob- ject grasping

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.710622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.863718Z digest=sha256:9f21fe26b8570fec94100f13c7c30196a892762cca1e8d0d2160531d2c7948b9

Observation 484ba86f-fe38-42fb-90fb-bde2230ef1e6 · outbound

This paper cites Planning optimal grasps.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Planning optimal grasps

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.556574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:25.961573Z digest=sha256:5d54b10e2db27b872ce19785be65a2c5dd0f4198adb719be76f0b99aea253706

Observation 574b1704-c839-4331-b855-a0d85ce8d7be · outbound

This paper cites Measurement of areas on a sphere using fibonacci and latitude–longitude lattices.Mathematical geo- sciences, 42:49–64, 2010.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Measurement of areas on a sphere using fibonacci and latitude–longitude lattices.Mathematical geo- sciences, 42:49–64, 2010

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.368086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:26.119442Z digest=sha256:14931959ec36a393a41bb9b68b33018fb46fe1a82f6da33c23762da65d6bd615

Observation 328d0923-740b-4194-9452-b0f4a504a948 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:40.213467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:26.238716Z digest=sha256:b102b4d23fc7c3460179e5f487676406a9cc28a025ffb1733f67b01799395ab6

Observation 12912eb0-d9b3-4aeb-ba45-eb7d14097568 · outbound

This paper cites Dexfuncgrasp: A robotic dexterous functional grasp dataset constructed from a cost-effective real-simulation annotation system.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexfuncgrasp: A robotic dexterous functional grasp dataset constructed from a cost-effective real-simulation annotation system

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.060031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:26.334294Z digest=sha256:0453014cf565fbcd1dce9be5549a7f84030eac1349b726775706669050d26772

Observation ad6eb191-2cd4-4300-a171-0bd7ed6e3f45 · outbound

This paper cites Tracking Objects with 3D Representation from Videos.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Tracking Objects with 3D Representation from Videos

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:34.283869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:26.478690Z digest=sha256:fe5ecb20cbfe7a7226c3b62c4cbdca5c40e6f80d4eca3dfd609893c8a7636628

Observation 6b743346-c9cd-4e7f-812a-6e861edd5c70 · outbound

This paper cites ManifoldPlus: A Robust and Scalable Watertight Manifold Surface Generation Method for Triangle Soups.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale ManifoldPlus: A Robust and Scalable Watertight Manifold Surface Generation Method for Triangle Soups

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.596066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.596066Z digest=sha256:377e0af852ae59de7c597acba50e7718e9b7d90586e8c30fdb4841a717212b7a

Observation 5e2de85e-1357-489e-b81c-0b4be878dc6c · outbound

This paper cites FunGrasp: Functional Grasping for Diverse Dexterous Hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale FunGrasp: Functional Grasping for Diverse Dexterous Hands

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:33.942387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:26.745663Z digest=sha256:6da36b1fa90eb7a8eaff3dcb1b591cea0008832ae560deefb02b32f245df69f5

Observation 92620677-bbac-42f9-97f0-a5ecd2e1e3a1 · outbound

This paper cites Omnispatial: Towards comprehensive spatial reasoning benchmark for vi- sion language models.arXiv preprint arXiv:2506.03135,.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Omnispatial: Towards comprehensive spatial reasoning benchmark for vi- sion language models.arXiv preprint arXiv:2506.03135,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.898492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.898492Z digest=sha256:7a59869ac71352bc418a63aa7de92c2146a8cdcaf74ed53a0829bc15230581dd

Observation c6b8a731-ecee-44c8-8c73-2ba20f0f1003 · outbound

This paper cites Hand-object contact consistency reasoning for human grasps generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Hand-object contact consistency reasoning for human grasps generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.959637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.959637Z digest=sha256:59136f61cb3d1724b521701b8b5eed67ca0f94d20bcca618c1b86f82ef4a683f

Observation 7e4f90b4-1b89-495e-a7cf-cc7f76242fce · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:27.086264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:27.086264Z digest=sha256:3ff97d22845542a17ef090b7ccbc792acaf12e6c69e8f9c3c6fabd541fdfec1c

Observation 1a0222a8-6ba0-4f43-b108-38d71125c67f · outbound

This paper cites Frogger: Fast robust grasp generation via the min-weight metric.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Frogger: Fast robust grasp generation via the min-weight metric

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.821992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.226058Z digest=sha256:df0d9151c9dbd4ef4426268eac7afe143b73bd97a52baed66a5cde86edfdbed4

Observation 31824d11-5009-4f46-a815-58f492c960c5 · outbound

This paper cites Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:27.315080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:27.315080Z digest=sha256:aa7bc852d95343b28fcfc2812ae9920a52557fd699d812092c82557c7f6828dc

Observation a546af55-d5d8-4155-85bf-900ce30e42e3 · outbound

This paper cites Semgrasp: Semantic grasp generation via language aligned discretization.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Semgrasp: Semantic grasp generation via language aligned discretization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.621484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.440427Z digest=sha256:229dbf133ddb511b851b944e52f58c767f71abc3887c7fc5e72d36559106fd39

Observation 759c2b4a-9570-4751-b51a-f58fee802b27 · outbound

This paper cites Incremental potential con- tact: intersection-and inversion-free, large-deformation dy- namics.ACM Trans.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Incremental potential con- tact: intersection-and inversion-free, large-deformation dy- namics.ACM Trans

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.456819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.640359Z digest=sha256:90a7d9df04188a1554da44a966486cd1b7105f8a09249eab39f15c9f0e1275a1

Observation 1a8d3641-35de-45f5-a327-e42838c22b24 · outbound

This paper cites Moka: Open-vocabulary robotic manipulation through mark-based visual prompting.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Moka: Open-vocabulary robotic manipulation through mark-based visual prompting

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.253008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.764261Z digest=sha256:96312e150c6530cc69df02e32d4686215d7ce543c9fcfd00948d20a9e66d6729

Observation 0b60d2f1-f597-4d13-ba6a-e553ca50149f · outbound

This paper cites Openshape: Scaling up 3d shape representation towards open-world understanding.Advances in neural information processing systems, 36:44860–44879, 2023.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Openshape: Scaling up 3d shape representation towards open-world understanding.Advances in neural information processing systems, 36:44860–44879, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.068759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.845319Z digest=sha256:485b94b5de52cd45485d735c2fed27f5e231bd9e6fbcdddd4dc704970f93d8d3

Observation 5c049be5-11c7-4ac1-82cc-b9b1553fe7d8 · outbound

This paper cites Partslip: Low-shot part segmentation for 3d point clouds via pretrained image- language models.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Partslip: Low-shot part segmentation for 3d point clouds via pretrained image- language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.921222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:27.989710Z digest=sha256:ec5277ba0e700b50acf3d4442dc2ceb74435fc02ae994c3f94929e66ecc682d5

Observation e68a3880-4e47-4baf-a13e-92e40d38dc7a · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:38.727799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:28.165184Z digest=sha256:3fd0f23641b3d6aa36010ca5e9216d0860530f536e5cb578ee3eda886c542b77

Observation fb3b7c14-7e54-4504-8893-65ac02a53e18 · outbound

This paper cites DexTrack: Towards Generalizable Neural Tracking Control for Dexterous Manipulation from Human References.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexTrack: Towards Generalizable Neural Tracking Control for Dexterous Manipulation from Human References

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:28.267833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:28.267833Z digest=sha256:9d85a31c8a79ae744c1bbc44390bc57b03adf26824a2229cd11b7e25e19855b8

Observation 0d78fade-e492-4176-b8b5-575977bf57e1 · outbound

This paper cites Cross-shape atten- tion for part segmentation of 3d point clouds.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Cross-shape atten- tion for part segmentation of 3d point clouds

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.553779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:28.417849Z digest=sha256:d4e5a8c78b453fc765304993693b5549cdfa162f1a23fe14a3ab23194850482f

Observation e4ff4f02-5337-40da-96a4-f85c842c4d7d · outbound

This paper cites Find Any Part in 3D.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Find Any Part in 3D

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:28.549836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:28.549836Z digest=sha256:4a5b4ca3bd26872f2970a9084faafd3d87e9fcd52d27528e2f0af99bc7ea4904

Observation 4fdee6b5-7f06-425c-b0a5-57c2ddd80e28 · outbound

This paper cites Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics, 2017.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics, 2017

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.415263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:28.696256Z digest=sha256:84315372328c6c80a30847bd9bd209d7ff44de0b1f222db7acf67f50c285d253

Observation 8ddbfee3-fb0e-4015-a812-3335b0616212 · outbound

This paper cites Isaac gym: High performance GPU based physics simulation for robot learning.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Isaac gym: High performance GPU based physics simulation for robot learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.302653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:28.841295Z digest=sha256:8a5a5fd692bb51a55deb6b344a217dff3b9396d21114e75819f8fee1ac94f612

Observation 74c17cf7-ce50-4338-b471-726231909b35 · outbound

This paper cites Introducing gpt-4o and more tools to chatgpt free users.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Introducing gpt-4o and more tools to chatgpt free users

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.183802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:28.935516Z digest=sha256:84a8488b1d92d6e9a7adcede98fbfed4c9b3582f0bfeedaf3bc9f6399c693df2

Observation 3e4b4fdf-03d6-48c6-86d5-e4618c81f0a5 · outbound

This paper cites Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.050894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:29.096366Z digest=sha256:ae6f90bedc1dd53b7761e040d233aad5e4b339e69b39ddf5df3fd386ede43a1a

Observation ec4a4cd4-a158-4452-93e3-611c56013000 · outbound

This paper cites Vpp: Efficient conditional 3d generation via voxel-point pro- gressive representation.Advances in Neural Information Processing Systems, 36:26744–26763, 2023.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Vpp: Efficient conditional 3d generation via voxel-point pro- gressive representation.Advances in Neural Information Processing Systems, 36:26744–26763, 2023

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.824946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:29.237495Z digest=sha256:77b72be6227c17e51ad4a07b2b6ac5d4d0914fb31c7bbf84f05ed81f99a23a14

Observation f57c9fbf-f3a9-4b58-a9ae-823e631aae83 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Shapellm: Universal 3d object understanding for embodied interaction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.403897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.403897Z digest=sha256:9c216f3d268136d0fcf97dcdcb76fde54ceb4207f67bd20d64d4354529faa542

Observation 001b863b-55dd-40ab-be74-f7f969b93e55 · outbound

This paper cites So- far: Language-grounded orientation bridges spatial reason- ing and object manipulation.CoRR, abs/2502.13143, 2025.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale So- far: Language-grounded orientation bridges spatial reason- ing and object manipulation.CoRR, abs/2502.13143, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.552214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.552214Z digest=sha256:d46e93c69ab79d43d2adf0a7827abb1f72f08b9d852f9f0786b1ac72d1ded095

Observation e24b2bae-4e0d-415f-a519-14d01369cc91 · outbound

This paper cites AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.763561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.763561Z digest=sha256:3fdcaa1d96d1343bd0958d5cef397d07f48215fa3e986ca3e84d79f18eaeb17d

Observation a39427a0-da84-4830-82de-88a5eaaede3b · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Learning transferable visual models from natural language supervi- sion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.927894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.927894Z digest=sha256:f94c8caa66228d2e7fdbd3c6eaa46a2dc60bbd95299213750823942f6e1ad0a4

Observation 6731e4e9-09df-439f-8968-8adb5322bf4a · outbound

This paper cites Curobo: Parallelized collision-free robot mo- tion generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Curobo: Parallelized collision-free robot mo- tion generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.071613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.071613Z digest=sha256:f08540b5898d0964e97780a4f88221aa10bad3c8bd93d3a831520eac04fc5aab

Observation 6e298da7-121a-49d0-8d11-301f2dd92c2b · outbound

This paper cites Segment Any Mesh.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Segment Any Mesh

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.129937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.129937Z digest=sha256:8189021b3f3c8bf2a5ef59c43b9c3cef808c13d48cfc53163a56d55ece02e171

Observation ae82df7e-2c73-43ce-b21a-f32634946c89 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Gemini: A Family of Highly Capable Multimodal Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.269101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.269101Z digest=sha256:5515d110b12ffcba44bda6986b8602795383bd0739e4d586156fde7545dcf3d4

Observation 5526c4d9-e073-42cb-9436-3053a2ba8d4d · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Octo: An Open-Source Generalist Robot Policy

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.354019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.354019Z digest=sha256:0e6f543248a92f050f7c738bf8c0359d7be1ef944d328fbd5b16101c38ebd2b1

Observation 40a46158-05c8-4132-b329-968a792b2634 · outbound

This paper cites Easy and fast evaluation of grasp stability by using ellipsoidal approx- imation of friction cone.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Easy and fast evaluation of grasp stability by using ellipsoidal approx- imation of friction cone

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.490954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:30.428822Z digest=sha256:540507b9c0502b3d908ad48b9379a9d153e8092c993b8adf7fd9596c585b7c7a

Observation ac00543b-ae4b-4794-a402-01a6edead64b · outbound

This paper cites Grasp’d: Differentiable contact-rich grasp syn- thesis for multi-fingered hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Grasp’d: Differentiable contact-rich grasp syn- thesis for multi-fingered hands

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.519370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.519370Z digest=sha256:1ec4f73231606310bd437266360f97f92c4534989f66fd45b64a472b3154372d

Observation 3ca1a846-5a69-4aaa-9be9-7219bd5dd0cf · outbound

This paper cites Fast-grasp’d: Dexterous multi- finger grasp generation through differentiable simulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Fast-grasp’d: Dexterous multi- finger grasp generation through differentiable simulation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.388564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:30.576277Z digest=sha256:3441b72d132163a0d9d503327b6fbd7f98ab536e708801975c50d4eea2aa3d36

Observation d8fbe70f-fd24-4389-b8ad-9e2ef53d9534 · outbound

This paper cites Unidexgrasp++: Im- proving dexterous grasping policy learning via geometry- aware curriculum and iterative generalist-specialist learning.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unidexgrasp++: Im- proving dexterous grasping policy learning via geometry- aware curriculum and iterative generalist-specialist learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.152665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:30.628617Z digest=sha256:3103d83fe346c3c92c03ecb45b801cccac7af77a29c5a77c311da059d6432281

Observation 554e3b7f-5dee-42a6-af31-bdc81f454ca8 · outbound

This paper cites Vlm see, robot do: Human demo video to robot action plan via vision language model.arXiv preprint arXiv:2410.08792, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Vlm see, robot do: Human demo video to robot action plan via vision language model.arXiv preprint arXiv:2410.08792, 2024

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.714416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.714416Z digest=sha256:55453d5653d15b9a5e97c52010869e96239b07e64ae58332ec4f6b13be1b0b79

Observation 4ac49ab8-ca17-416e-926e-5367312bc825 · outbound

This paper cites DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.777176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.777176Z digest=sha256:6add1b04e2e3cd5bb027f7b52febd5a27a4c227d357061743f3eb223170d193b

Observation 28018a36-fff2-4a7c-bd17-b74eef8cb7b9 · outbound

This paper cites Dexgraspnet: A large-scale robotic dexterous grasp dataset for general ob- jects based on simulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexgraspnet: A large-scale robotic dexterous grasp dataset for general ob- jects based on simulation

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.949010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:30.843470Z digest=sha256:8a7c98274e472c30eac91d40943de92de72e445b70d07c93b2a54984a79c41af

Observation 6c5e48df-43f6-4884-893b-7e377768460b · outbound

This paper cites Approx- imate convex decomposition for 3d meshes with collision- aware concavity and tree search.ACM Transactions on Graphics (TOG), 41(4):1–18, 2022.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Approx- imate convex decomposition for 3d meshes with collision- aware concavity and tree search.ACM Transactions on Graphics (TOG), 41(4):1–18, 2022

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.757987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:30.888914Z digest=sha256:673df0959ed7849845b7310bbd72b3e6e90dc6800ab526963b3d643ce672e69f

Observation 1f2c84d2-8bce-4fd9-9a4f-1b47cdc73c5d · outbound

This paper cites Grasp as You Say: Language-guided Dexterous Grasp Generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Grasp as You Say: Language-guided Dexterous Grasp Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.979171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.979171Z digest=sha256:26c202ca437693709f40d0524fcd664e857546d38c590b195561493aed3737d9

Observation 1b8bc759-778b-4857-974b-eebede907b71 · outbound

This paper cites Cross- category functional grasp transfer.IEEE Robotics and Au- tomation Letters, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Cross- category functional grasp transfer.IEEE Robotics and Au- tomation Letters, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.679576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.042231Z digest=sha256:2344a72a0914c1e8c26ad42ecca96191e850e8b1177f4fe11f2597e50c5950cf

Observation 9b1cd857-1350-422b-8ae3-e6bb820692c4 · outbound

This paper cites Florence-2: Advancing a unified representation for a variety of vision tasks.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Florence-2: Advancing a unified representation for a variety of vision tasks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.435931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.102850Z digest=sha256:d967e60920ac40b986f178fb5399fa3aa2eeb75b48e6e57e9f0dda4e21f79b37

Observation e8062937-6c91-4d2a-8726-304fa3ac0de8 · outbound

This paper cites Dexterous grasp transformer, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexterous grasp transformer, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.236200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.225943Z digest=sha256:653de54ef22c26d1dd944f3b715dff3b87896fa200b08b06949eac546cb447af

Observation 625474bd-9997-49bd-8989-64859e589787 · outbound

This paper cites Unidexgrasp: Universal robotic dexterous grasping via learning diverse proposal generation and goal-conditioned policy.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unidexgrasp: Universal robotic dexterous grasping via learning diverse proposal generation and goal-conditioned policy

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.029771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.330212Z digest=sha256:98ad79694d71dd22b4ccb1c1bf59756ee4d7562cf5208b154659d8f3869725dc

Observation 291ccc3b-36d7-4dfb-b3d9-58a21e24e35b · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:31.477714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:31.477714Z digest=sha256:8c39277e049cff5e807799c289229716e0893e83702041d8dbc0a1302a3cd28a

Observation 64cbfbd5-03ec-4f53-863f-3a8d923d1356 · outbound

This paper cites Oakink: A large-scale knowledge repos- itory for understanding hand-object interaction.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Oakink: A large-scale knowledge repos- itory for understanding hand-object interaction

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.833688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.600400Z digest=sha256:75422b84f5881f9e18464c7e97c8e40e07b35709829ef64ae527d0371eec3d11

Observation fa30e64b-9fce-4439-a8c9-2e36a3b99977 · outbound

This paper cites SAMPart3D: Segment Any Part in 3D Objects.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale SAMPart3D: Segment Any Part in 3D Objects

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:31.778163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:31.778163Z digest=sha256:d66a26e955dca63036eb1744d5acf6d715b5ddd6f334c8a5c99c9fba6deda350

Observation e58fd632-e8c1-4189-bd3e-b23df31690bb · outbound

This paper cites Graspxl: Generating grasping motions for di- verse objects at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Graspxl: Generating grasping motions for di- verse objects at scale

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.618900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:31.977479Z digest=sha256:b81bd335b6fd607d01b8fd737ff8749939e5f64d52687da0ebb899df88ec753c

Observation 3722ba1b-55d3-4e96-a9a9-5a5a7b32dfdc · outbound

This paper cites Dexgrasp- net 2.0: Learning generative dexterous grasping in large- scale synthetic cluttered scenes.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexgrasp- net 2.0: Learning generative dexterous grasping in large- scale synthetic cluttered scenes

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.413575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:32.179501Z digest=sha256:e260d1e89b0bb8d02d1964ed5d26c2eb1dbcb9950e395955152454a1933f5c60

Observation 905ff705-9adf-4d0a-a0f1-1fa1ea1380e2 · outbound

This paper cites DexGrasp-Diffusion: Diffusion-based Unified Functional Grasp Synthesis Method for Multi-Dexterous Robotic Hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexGrasp-Diffusion: Diffusion-based Unified Functional Grasp Synthesis Method for Multi-Dexterous Robotic Hands

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:33.373364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:32.349340Z digest=sha256:a7d425119cfa5af3ec9d625fd241d249836d7c625f1e5f2327dbbb431d6c7a2b

Observation 9bcd0b44-b774-453c-b81f-6bc6f688c1c5 · outbound

This paper cites Point transformer.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Point transformer

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:32.546464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:32.546464Z digest=sha256:b856fa4235f6a280955de90a953dba695d04c84603e37d8acad8f3de7fee6e60

Observation fb1b47f1-2cd9-4ef9-8f93-15d57347892a · outbound

This paper cites Transfusion: Pre- dict the next token and diffuse images with one multi- modal model.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Transfusion: Pre- dict the next token and diffuse images with one multi- modal model

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.139936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:32.668921Z digest=sha256:34e803bbb62df9f0fe5fa3ee34e5ce8ec1cd75acbc7273f7ac36f9326ecb6907

Observation 98c75145-325c-4597-b82a-c9d1b8c10234 · outbound

This paper cites Uni3d: Exploring uni- fied 3d representation at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Uni3d: Exploring uni- fied 3d representation at scale

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:34.832315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:32.831965Z digest=sha256:b01ba2b0002048125bdabe8abd1614261e80402f7b0d37ff7649d995496a72e9

Observation 691352d1-c632-4cef-8c1a-a8e3adc332f2 · outbound

This paper cites PartSLIP++: Enhancing Low-Shot 3D Part Segmentation via Multi-View Instance Segmentation and Maximum Likelihood Estimation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale PartSLIP++: Enhancing Low-Shot 3D Part Segmentation via Multi-View Instance Segmentation and Maximum Likelihood Estimation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:32.980977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:32.980977Z digest=sha256:5b4c3e79554421ed5baee29c536fd98f0792a9aff843eb044676b1171ab1efa8

Observation 02869e3b-8dfb-4bfb-8f96-8e25208f20f0 · outbound

This paper cites embedded inside.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale embedded inside

Reference 73

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:29:34.596928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-06T20:29:33.097470Z digest=sha256:d2b0af8ce3e5b0b340f3a7ba3b63ec44f2b2cebcf67942a861eb7dca5c656950

Pith citing papers

Observation 8a4016a4-f65d-497a-8bd7-098b8b8eb7f8 · inbound

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge cites this paper.

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:42:41.415484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T15:42:41.363422Z digest=sha256:e53b7c6684f289f8e5031593ac86608cdbcbeaff4820eb5ca33d4dc39182f308

Observation 9c96aaab-977f-4a81-9243-af342b15fcf0 · inbound

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos cites this paper.

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:33:39.445246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:33:39.445246Z digest=sha256:b0283c098780b761e3f672348f954fd3eb1c2f7b6ca03e8013bd394b12d94244

Observation 71f0ccd7-34d9-47a4-8cd1-8c130286bba1 · inbound

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands cites this paper.

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:52:38.724758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T13:52:35.939944Z digest=sha256:568d8e3b784cea18f626e269e93cdf931f8ecd5d19109b8c58869d7702630186

Observation efb12478-110b-431a-ab25-204dd742e19b · inbound

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models cites this paper.

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:02:24.955096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T06:01:30.803128Z digest=sha256:2d8d8e5ea9a58a6663b352b9e7f4d49b974fb335d24cdbeff0265c585d94df19

Observation 47baa568-4a1b-4428-bc15-319925288dd0 · inbound

BiDexGrasp: Coordinated Bimanual Dexterous Grasps across Object Geometries and Sizes cites this paper.

BiDexGrasp: Coordinated Bimanual Dexterous Grasps across Object Geometries and Sizes DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:56.957350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T18:49:49.433772Z digest=sha256:bbb3812a899a91c532c884b01104cf0a65b3ad98e4b962a5f545c9f92fb287e8

Observation 5f05e673-cf52-480f-bb0b-9bc2bc9af562 · inbound

BLaDA: Bridging Language to Functional Dexterous Actions within 3DGS Fields cites this paper.

BLaDA: Bridging Language to Functional Dexterous Actions within 3DGS Fields DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:51:25.793833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T17:25:09.012355Z digest=sha256:24658e25765734248e6ecca164e8e37aa81eec07829dfcfe2181fb28e29d0882

Observation a7a90edd-ed5b-40ad-bdd1-bc8a9ed4a4b1 · inbound

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation cites this paper.

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T06:04:34.677454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-08T05:55:32.387354Z digest=sha256:3a94bb52f3a41bf42c6593fd6a50bb59a7a6db00363650770992c3d4ef0fb75d

Observation 798058ca-a00a-4edb-b787-4ba0a6250bdb · inbound

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation cites this paper.

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T16:02:59.358284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:02:59.358284Z digest=sha256:06dce59840474dbe3a423a82064a0d3ab36ca766480eca9171dc2255371f84c9