Pith. sign in

Paper Citation Record · LEDGER

DexVLG: Dexterous Vision-Language-Grasp Model at Scale

As of 11 August 2026, this Paper Citation Record lists 73 of 73 outbound references and 8 inbound Pith citation observations for arXiv:2507.02747.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02747 v1

Coverage vector

measured 73 of 73 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:29:33.097470Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:33:39.445246Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T06:04:34.676226Z

Reference resolution

73 of 73 outbound references displayed

  • verified exact3
  • verified fuzzy35
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 10f0d90b-4765-4165-b495-0467919d6626 · outbound

This paper cites GPT-4 Technical Report.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:23.959143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:23.959143Z digest=sha256:da1e45c612d3e336b6950466d666f2ec5855791197c5a784279a4544376b3d6c

Observation 0c4f6461-ba29-48e5-9111-f06bfd3539c9 · outbound

This paper cites Dexterous functional grasping.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexterous functional grasping

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.919023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:24.089645Z digest=sha256:e4c0fea94ab89e208d09b0430b21a3b42800fa1de5e4160f240cd3fe0e7e94a1

Observation ede89f73-b332-4bb4-b453-7dec4abab9d5 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.283858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.283858Z digest=sha256:9e77abd588b76496ac38c19d3d18dd16ae3ea61c4b1420edb5100bdde4b53092

Observation dfcc98e9-132b-4609-a7d5-71429ee5c4e2 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.467366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.467366Z digest=sha256:89f6da158e658e2e17c86d3da7d8225a9d772eab628467281de74e68c0e782c6

Observation a15c7294-a4df-49b9-ac51-e115657935ca · outbound

This paper cites Task-Oriented Dexterous Hand Pose Synthesis Using Differentiable Grasp Wrench Boundary Estimator.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Task-Oriented Dexterous Hand Pose Synthesis Using Differentiable Grasp Wrench Boundary Estimator

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.623262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.623262Z digest=sha256:a1d27ac2e1799fb6ba8aac996117c8405820f4f4f8c552f0ff4bc4d15e04628e

Observation 60e7b811-93d2-43be-ad01-0fa11a225596 · outbound

This paper cites BODex: Scalable and Efficient Robotic Dexterous Grasp Synthesis Using Bilevel Optimization.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale BODex: Scalable and Efficient Robotic Dexterous Grasp Synthesis Using Bilevel Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.749825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.749825Z digest=sha256:15a48f2516762a95ffc3f6cbc136de7961274b41a5d25ecd3d7aba21cc9bf849

Observation 83ecaee0-2962-41ff-8a05-b3960414cf34 · outbound

This paper cites SpringGrasp: Synthesizing Compliant, Dexterous Grasps under Shape Uncertainty.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale SpringGrasp: Synthesizing Compliant, Dexterous Grasps under Shape Uncertainty

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:24.877321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:24.877321Z digest=sha256:21777991e0a44d8663d046e7a845368baa4ebc1acdba80e9fbb24c562f2d5b42

Observation 60e8948a-ac7d-47c4-b74d-5e1f47b1063a · outbound

This paper cites Learning Robust Real-World Dexterous Grasping Policies via Implicit Shape Augmentation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Learning Robust Real-World Dexterous Grasping Policies via Implicit Shape Augmentation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:25.009478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:25.009478Z digest=sha256:4fb6328b764ff197faed2b664a410093c1d845cec7a10a50208f17200360fd63

Observation 11dfaaf7-284c-492b-9360-c96c594889e4 · outbound

This paper cites Syn- thesis and optimization of force closure grasps via sequential semidefinite programming.Robotics Research: Volume 1, pages 285–305, 2018.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Syn- thesis and optimization of force closure grasps via sequential semidefinite programming.Robotics Research: Volume 1, pages 285–305, 2018

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.786579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.178223Z digest=sha256:0e85f82670e3b404a9cf80f35d755652f0d05abcc1fe7e8d5f285f7d698ac7c1

Observation c8bb98b4-d7fe-44a9-9a6c-e77b02dd11b5 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:41.670671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.316048Z digest=sha256:04550c007cab2bd4d16a615e811d566b266da65690a4ff1cda7dc6849a21db7a

Observation be662dab-1b63-43c1-98ea-fff45681b87e · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Objaverse: A universe of annotated 3d objects

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.529274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.408782Z digest=sha256:864d2c7045b0daba2b6ff34eb6fafa5fd533bc4fdebb13da0620e473e1140e56

Observation 11c6b362-0cd5-40d8-a3af-8de5df796ed0 · outbound

This paper cites Open6dor: Benchmarking open-instruction 6-dof object rearrangement and a vlm-based approach.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Open6dor: Benchmarking open-instruction 6-dof object rearrangement and a vlm-based approach

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:41.246754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.517974Z digest=sha256:24457ef3ac80a10674314fb822a80a5a87cdaef905ce8e342f07fd6797acb13f

Observation ee315b7d-9b6b-46ff-a9b3-46f4cd2896b2 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:40.986668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.587196Z digest=sha256:5fc4aa5d19a6bc403156b6764952de9acab8c51654f08d97e58087d4d698a4a9

Observation 5a7b652c-04d2-4b32-abec-e36d42a1700b · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale An image is worth 16x16 words: Transformers for image recognition at scale

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.865746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.708066Z digest=sha256:f35adc6f691a34a3ad21152899e19322fe9bbe59c9934aecfe8b79846cd39f0b

Observation 9129cffc-20ea-4882-8833-ecb64fac30c9 · outbound

This paper cites Graspnet-1billion: A large-scale benchmark for general ob- ject grasping.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Graspnet-1billion: A large-scale benchmark for general ob- ject grasping

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.710622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.863718Z digest=sha256:e7e505ddbdc52f0dd1103286068c3bc980291050f5cde5fb60519d4e903bde41

Observation 484ba86f-fe38-42fb-90fb-bde2230ef1e6 · outbound

This paper cites Planning optimal grasps.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Planning optimal grasps

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.556574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:25.961573Z digest=sha256:35400abf85b1182c9d69d8f0e6c45a417b13ed3e9bfd55f53c1a0399d30a6283

Observation 574b1704-c839-4331-b855-a0d85ce8d7be · outbound

This paper cites Measurement of areas on a sphere using fibonacci and latitude–longitude lattices.Mathematical geo- sciences, 42:49–64, 2010.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Measurement of areas on a sphere using fibonacci and latitude–longitude lattices.Mathematical geo- sciences, 42:49–64, 2010

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.368086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:26.119442Z digest=sha256:34afb4e380d060c34151a30906cdf4c642a9ef69381ee7fae2e8bf47e3207aa3

Observation 328d0923-740b-4194-9452-b0f4a504a948 · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:40.213467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:26.238716Z digest=sha256:a1f137a8b97972205da3da34eec329582632f6e126be472429fa5c62e4cc2181

Observation 12912eb0-d9b3-4aeb-ba45-eb7d14097568 · outbound

This paper cites Dexfuncgrasp: A robotic dexterous functional grasp dataset constructed from a cost-effective real-simulation annotation system.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexfuncgrasp: A robotic dexterous functional grasp dataset constructed from a cost-effective real-simulation annotation system

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:40.060031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:26.334294Z digest=sha256:e8ad90bdd6cea3ed73b7596bf6e658e93a7bed14e781869f2edfb7942b5bc0a1

Observation ad6eb191-2cd4-4300-a171-0bd7ed6e3f45 · outbound

This paper cites Tracking Objects with 3D Representation from Videos.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Tracking Objects with 3D Representation from Videos

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:34.283869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:26.478690Z digest=sha256:400d0574cb4019108445261f23780334b3aa886b5eec7e1a3e480890fc03609e

Observation 6b743346-c9cd-4e7f-812a-6e861edd5c70 · outbound

This paper cites ManifoldPlus: A Robust and Scalable Watertight Manifold Surface Generation Method for Triangle Soups.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale ManifoldPlus: A Robust and Scalable Watertight Manifold Surface Generation Method for Triangle Soups

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.596066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.596066Z digest=sha256:fbad6532bee55a49f950d044812a65c41bbb2fde1360c64f02252dd3ef185bde

Observation 5e2de85e-1357-489e-b81c-0b4be878dc6c · outbound

This paper cites FunGrasp: Functional Grasping for Diverse Dexterous Hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale FunGrasp: Functional Grasping for Diverse Dexterous Hands

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:33.942387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:26.745663Z digest=sha256:20cadfb30026b9aa882de3b194fb1dd1fbbbd04070ec38823233d6ed7484a9b2

Observation 92620677-bbac-42f9-97f0-a5ecd2e1e3a1 · outbound

This paper cites Omnispatial: Towards comprehensive spatial reasoning benchmark for vi- sion language models.arXiv preprint arXiv:2506.03135,.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Omnispatial: Towards comprehensive spatial reasoning benchmark for vi- sion language models.arXiv preprint arXiv:2506.03135,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.898492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.898492Z digest=sha256:7e6230ff671ae55b73c245c5cf1266019d3180e812dcebc668de7e8a7435a976

Observation c6b8a731-ecee-44c8-8c73-2ba20f0f1003 · outbound

This paper cites Hand-object contact consistency reasoning for human grasps generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Hand-object contact consistency reasoning for human grasps generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:26.959637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:26.959637Z digest=sha256:a49e24d3af93870cffcfadc70273994cda7d7ae48f15809ac62f05b5c6a29e20

Observation 7e4f90b4-1b89-495e-a7cf-cc7f76242fce · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:27.086264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:27.086264Z digest=sha256:dc324e22c778c4762ea9379211862e9e69646d5f50e75f25786d43d59306f093

Observation 1a0222a8-6ba0-4f43-b108-38d71125c67f · outbound

This paper cites Frogger: Fast robust grasp generation via the min-weight metric.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Frogger: Fast robust grasp generation via the min-weight metric

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.821992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.226058Z digest=sha256:6613be7227cf93575c82af19b9f96945ca2d856df3fc3b009a4a4c053e5143bf

Observation 31824d11-5009-4f46-a815-58f492c960c5 · outbound

This paper cites Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:27.315080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:27.315080Z digest=sha256:8239389e6b2876a563a5761c5d8e65f1f54ccb2131401d6d831b1eec943f072a

Observation a546af55-d5d8-4155-85bf-900ce30e42e3 · outbound

This paper cites Semgrasp: Semantic grasp generation via language aligned discretization.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Semgrasp: Semantic grasp generation via language aligned discretization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.621484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.440427Z digest=sha256:2cf3a1ea34c01e754a51079b3e1ea82fe8c24092d85114b726093b6bff8c6020

Observation 759c2b4a-9570-4751-b51a-f58fee802b27 · outbound

This paper cites Incremental potential con- tact: intersection-and inversion-free, large-deformation dy- namics.ACM Trans.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Incremental potential con- tact: intersection-and inversion-free, large-deformation dy- namics.ACM Trans

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.456819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.640359Z digest=sha256:dfe37e0261d1c46f13e5dbf635eeed8358996a0989472dc5fbd4bb02afdddb1c

Observation 1a8d3641-35de-45f5-a327-e42838c22b24 · outbound

This paper cites Moka: Open-vocabulary robotic manipulation through mark-based visual prompting.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Moka: Open-vocabulary robotic manipulation through mark-based visual prompting

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.253008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.764261Z digest=sha256:ebea6bc3db0569384b673cb93ff2acea4c5da9fb552296858b6f13bb68dfaa76

Observation 0b60d2f1-f597-4d13-ba6a-e553ca50149f · outbound

This paper cites Openshape: Scaling up 3d shape representation towards open-world understanding.Advances in neural information processing systems, 36:44860–44879, 2023.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Openshape: Scaling up 3d shape representation towards open-world understanding.Advances in neural information processing systems, 36:44860–44879, 2023

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:39.068759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.845319Z digest=sha256:078de40c0d276ce2d753151f113b31c5fbcdbb90d4bb5d0d97333ac806687686

Observation 5c049be5-11c7-4ac1-82cc-b9b1553fe7d8 · outbound

This paper cites Partslip: Low-shot part segmentation for 3d point clouds via pretrained image- language models.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Partslip: Low-shot part segmentation for 3d point clouds via pretrained image- language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.921222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:27.989710Z digest=sha256:7abbc1bcc231e1ca2facb7d7c06ffe49fc64e9e0a027c4732e3ab87a14658056

Observation e68a3880-4e47-4baf-a13e-92e40d38dc7a · outbound

This paper cites an unresolved cited work.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:29:38.727799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:28.165184Z digest=sha256:99165e3fc38a2150f1ce2ed8191027314c0494a1fcc6fc37a9f82fcf6c44b704

Observation fb3b7c14-7e54-4504-8893-65ac02a53e18 · outbound

This paper cites DexTrack: Towards Generalizable Neural Tracking Control for Dexterous Manipulation from Human References.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexTrack: Towards Generalizable Neural Tracking Control for Dexterous Manipulation from Human References

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:28.267833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:28.267833Z digest=sha256:7228bd6b44d17abdbefa1ec6008292d27fe1de463d6b367135df4ff3ab97b9a9

Observation 0d78fade-e492-4176-b8b5-575977bf57e1 · outbound

This paper cites Cross-shape atten- tion for part segmentation of 3d point clouds.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Cross-shape atten- tion for part segmentation of 3d point clouds

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.553779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:28.417849Z digest=sha256:cfc63813c4a2bda37c942d34090a5cd7af646de80856dda264ee2a5504250397

Observation e4ff4f02-5337-40da-96a4-f85c842c4d7d · outbound

This paper cites Find Any Part in 3D.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Find Any Part in 3D

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:28.549836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:28.549836Z digest=sha256:6e0c3f8729f870b28e9c376ee8831a55e2464096e12dfafc8bf9b5d9f12e6f50

Observation 4fdee6b5-7f06-425c-b0a5-57c2ddd80e28 · outbound

This paper cites Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics, 2017.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dex-net 2.0: Deep learning to plan robust grasps with synthetic point clouds and analytic grasp metrics, 2017

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.415263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:28.696256Z digest=sha256:d1c57f55e5bbc038f0d149bcf705926d7aa4fd7224598bca15b2807d3a74c3e4

Observation 8ddbfee3-fb0e-4015-a812-3335b0616212 · outbound

This paper cites Isaac gym: High performance GPU based physics simulation for robot learning.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Isaac gym: High performance GPU based physics simulation for robot learning

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.302653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:28.841295Z digest=sha256:9ed0a9217f38d6caae91de91ce92027266ea2294bdb9e925341f975868ca4b72

Observation 74c17cf7-ce50-4338-b471-726231909b35 · outbound

This paper cites Introducing gpt-4o and more tools to chatgpt free users.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Introducing gpt-4o and more tools to chatgpt free users

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.183802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:28.935516Z digest=sha256:772f6dd788e751229081106e35cf6a1ad9201011d9e516c27d59fc8fd039e7e7

Observation 3e4b4fdf-03d6-48c6-86d5-e4618c81f0a5 · outbound

This paper cites Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:38.050894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:29.096366Z digest=sha256:f42ecc00e5bff20780ccf9b42eeba77a0efd76cb848fdf40f08c892574cef661

Observation ec4a4cd4-a158-4452-93e3-611c56013000 · outbound

This paper cites Vpp: Efficient conditional 3d generation via voxel-point pro- gressive representation.Advances in Neural Information Processing Systems, 36:26744–26763, 2023.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Vpp: Efficient conditional 3d generation via voxel-point pro- gressive representation.Advances in Neural Information Processing Systems, 36:26744–26763, 2023

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.824946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:29.237495Z digest=sha256:807bf20e8abcf97c328c17f24a8b608e222502b2d2097c36044795827f3afe37

Observation f57c9fbf-f3a9-4b58-a9ae-823e631aae83 · outbound

This paper cites Shapellm: Universal 3d object understanding for embodied interaction.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Shapellm: Universal 3d object understanding for embodied interaction

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.403897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.403897Z digest=sha256:4b35880f46caed9d56c9f370a0bf5cd0e257f4f467c78f1b74ad8f8774f2b343

Observation 001b863b-55dd-40ab-be74-f7f969b93e55 · outbound

This paper cites So- far: Language-grounded orientation bridges spatial reason- ing and object manipulation.CoRR, abs/2502.13143, 2025.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale So- far: Language-grounded orientation bridges spatial reason- ing and object manipulation.CoRR, abs/2502.13143, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.552214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.552214Z digest=sha256:53e7501e81437e234b6b8a0d6bf462134417b46ae133a66c15fb481f6d6bb5c5

Observation e24b2bae-4e0d-415f-a519-14d01369cc91 · outbound

This paper cites AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale AnyTeleop: A General Vision-Based Dexterous Robot Arm-Hand Teleoperation System

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.763561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.763561Z digest=sha256:c091b9b24616144b3ee7df8b7697034938e2dc3b6d8049f2b796f64bf871ed4d

Observation a39427a0-da84-4830-82de-88a5eaaede3b · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Learning transferable visual models from natural language supervi- sion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:29.927894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:29.927894Z digest=sha256:1ce5c3cd4bdd785a9fa073013e0b84da5f0d7c0887eaf59a108f52ce2ab8526b

Observation 6731e4e9-09df-439f-8968-8adb5322bf4a · outbound

This paper cites Curobo: Parallelized collision-free robot mo- tion generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Curobo: Parallelized collision-free robot mo- tion generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.071613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.071613Z digest=sha256:8ba010d89a566af24ee5d21e0bdf6e3932849a8be3943e9444f0ce325671eaa9

Observation 6e298da7-121a-49d0-8d11-301f2dd92c2b · outbound

This paper cites Segment Any Mesh.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Segment Any Mesh

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.129937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.129937Z digest=sha256:265dbd0acc85912eab3dc408d2c9191890bffb34ed90561afee55f696494140c

Observation ae82df7e-2c73-43ce-b21a-f32634946c89 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Gemini: A Family of Highly Capable Multimodal Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.269101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.269101Z digest=sha256:04bfc18a6866032773bfab300610a7ef47abafb6ea990477a3933e3fe7a54446

Observation 5526c4d9-e073-42cb-9436-3053a2ba8d4d · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Octo: An Open-Source Generalist Robot Policy

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.354019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.354019Z digest=sha256:aa62e992fcbb0e0fa87232b7c213fbedb2d78efde58209b5b64fe8ad17bb0fec

Observation 40a46158-05c8-4132-b329-968a792b2634 · outbound

This paper cites Easy and fast evaluation of grasp stability by using ellipsoidal approx- imation of friction cone.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Easy and fast evaluation of grasp stability by using ellipsoidal approx- imation of friction cone

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.490954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:30.428822Z digest=sha256:5fcb43bc896dcc5e26e4eb3dce5271d65f92433ebea6d193546b6e0f15f287ec

Observation ac00543b-ae4b-4794-a402-01a6edead64b · outbound

This paper cites Grasp’d: Differentiable contact-rich grasp syn- thesis for multi-fingered hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Grasp’d: Differentiable contact-rich grasp syn- thesis for multi-fingered hands

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.519370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.519370Z digest=sha256:0257e50bd00956350af98dd02bf7ac7a6fe8d57bd4508f318492652241b3f473

Observation 3ca1a846-5a69-4aaa-9be9-7219bd5dd0cf · outbound

This paper cites Fast-grasp’d: Dexterous multi- finger grasp generation through differentiable simulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Fast-grasp’d: Dexterous multi- finger grasp generation through differentiable simulation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.388564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:30.576277Z digest=sha256:2a65441bc636a92fb2b1a3d326486bdc4533508c2f2b6c5bb3f6adaf3a98b5a2

Observation d8fbe70f-fd24-4389-b8ad-9e2ef53d9534 · outbound

This paper cites Unidexgrasp++: Im- proving dexterous grasping policy learning via geometry- aware curriculum and iterative generalist-specialist learning.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unidexgrasp++: Im- proving dexterous grasping policy learning via geometry- aware curriculum and iterative generalist-specialist learning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:37.152665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:30.628617Z digest=sha256:c18933ec94c57e3afc1b332adea906ca78b1bc16ec782f11901f3df74e3ebdb0

Observation 554e3b7f-5dee-42a6-af31-bdc81f454ca8 · outbound

This paper cites Vlm see, robot do: Human demo video to robot action plan via vision language model.arXiv preprint arXiv:2410.08792, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Vlm see, robot do: Human demo video to robot action plan via vision language model.arXiv preprint arXiv:2410.08792, 2024

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.714416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.714416Z digest=sha256:4085bb64549a96c551e1456ca386f83eb3405d8cadacb5b5d5be601cc50e9a2a

Observation 4ac49ab8-ca17-416e-926e-5367312bc825 · outbound

This paper cites DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexCap: Scalable and Portable Mocap Data Collection System for Dexterous Manipulation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.777176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.777176Z digest=sha256:300a43070f78bd518f24265c887b25876d96cd366965f254bf4dd85788fac8ce

Observation 28018a36-fff2-4a7c-bd17-b74eef8cb7b9 · outbound

This paper cites Dexgraspnet: A large-scale robotic dexterous grasp dataset for general ob- jects based on simulation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexgraspnet: A large-scale robotic dexterous grasp dataset for general ob- jects based on simulation

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.949010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:30.843470Z digest=sha256:e7611e67c19e07360e42efd0b7994f8c77a56f46db7fc555d1e595eaa67ba392

Observation 6c5e48df-43f6-4884-893b-7e377768460b · outbound

This paper cites Approx- imate convex decomposition for 3d meshes with collision- aware concavity and tree search.ACM Transactions on Graphics (TOG), 41(4):1–18, 2022.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Approx- imate convex decomposition for 3d meshes with collision- aware concavity and tree search.ACM Transactions on Graphics (TOG), 41(4):1–18, 2022

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.757987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:30.888914Z digest=sha256:51f35c94ab812d1bb411ad30b077b03b57159fb94f9b627238b8d1f95b3488cb

Observation 1f2c84d2-8bce-4fd9-9a4f-1b47cdc73c5d · outbound

This paper cites Grasp as You Say: Language-guided Dexterous Grasp Generation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Grasp as You Say: Language-guided Dexterous Grasp Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:30.979171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:30.979171Z digest=sha256:3151d867fe3ca00eb2ba65b78957729420bf955e79c050dfda55f0fda24ba5a3

Observation 1b8bc759-778b-4857-974b-eebede907b71 · outbound

This paper cites Cross- category functional grasp transfer.IEEE Robotics and Au- tomation Letters, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Cross- category functional grasp transfer.IEEE Robotics and Au- tomation Letters, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.679576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.042231Z digest=sha256:d751f71b93c33e469397f662957035810cceba2fa46315950ac84533803fcd3c

Observation 9b1cd857-1350-422b-8ae3-e6bb820692c4 · outbound

This paper cites Florence-2: Advancing a unified representation for a variety of vision tasks.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Florence-2: Advancing a unified representation for a variety of vision tasks

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.435931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.102850Z digest=sha256:c3603757226d9fec8b31c7601eb1e41560719c421c6fa675afbfb54df0200fc1

Observation e8062937-6c91-4d2a-8726-304fa3ac0de8 · outbound

This paper cites Dexterous grasp transformer, 2024.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexterous grasp transformer, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.236200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.225943Z digest=sha256:b6eb474acdfe19d5b92634d351097668d3d632909ec68ea3b10ad086a50716e7

Observation 625474bd-9997-49bd-8989-64859e589787 · outbound

This paper cites Unidexgrasp: Universal robotic dexterous grasping via learning diverse proposal generation and goal-conditioned policy.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Unidexgrasp: Universal robotic dexterous grasping via learning diverse proposal generation and goal-conditioned policy

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:36.029771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.330212Z digest=sha256:0f9514c70195fded415c185a36faba3e7d89e65ca45de99bc92775c3812c4372

Observation 291ccc3b-36d7-4dfb-b3d9-58a21e24e35b · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:31.477714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:31.477714Z digest=sha256:8a0ac79535c12832b1d72490449468992d97013ad79f292af7b798c95f9ff754

Observation 64cbfbd5-03ec-4f53-863f-3a8d923d1356 · outbound

This paper cites Oakink: A large-scale knowledge repos- itory for understanding hand-object interaction.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Oakink: A large-scale knowledge repos- itory for understanding hand-object interaction

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.833688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.600400Z digest=sha256:d56e0335f4b75150e758ec07d809cd0d5454d09db7c6e034c265d994e7013a4a

Observation fa30e64b-9fce-4439-a8c9-2e36a3b99977 · outbound

This paper cites SAMPart3D: Segment Any Part in 3D Objects.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale SAMPart3D: Segment Any Part in 3D Objects

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:31.778163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:31.778163Z digest=sha256:612ea3bdb121cf72f25f7f90f6c48f2e80484ebdf5f15232cdd6df3addc52b7f

Observation e58fd632-e8c1-4189-bd3e-b23df31690bb · outbound

This paper cites Graspxl: Generating grasping motions for di- verse objects at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Graspxl: Generating grasping motions for di- verse objects at scale

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.618900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:31.977479Z digest=sha256:689b12fe6ed745530117186333abe67ef566d21c78d483b73fc07868432c9415

Observation 3722ba1b-55d3-4e96-a9a9-5a5a7b32dfdc · outbound

This paper cites Dexgrasp- net 2.0: Learning generative dexterous grasping in large- scale synthetic cluttered scenes.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Dexgrasp- net 2.0: Learning generative dexterous grasping in large- scale synthetic cluttered scenes

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.413575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:32.179501Z digest=sha256:ffc16b160688ccea072d0263b186fa7b7b88c1e007ac2297a362d8a49866f064

Observation 905ff705-9adf-4d0a-a0f1-1fa1ea1380e2 · outbound

This paper cites DexGrasp-Diffusion: Diffusion-based Unified Functional Grasp Synthesis Method for Multi-Dexterous Robotic Hands.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale DexGrasp-Diffusion: Diffusion-based Unified Functional Grasp Synthesis Method for Multi-Dexterous Robotic Hands

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T20:29:33.373364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:32.349340Z digest=sha256:d9b0360dd9b736002b96445f97cf52f493a9122cf88e4e394ed857f8e73afe9a

Observation 9bcd0b44-b774-453c-b81f-6bc6f688c1c5 · outbound

This paper cites Point transformer.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Point transformer

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:32.546464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:32.546464Z digest=sha256:c466b290bcf1bb2bb4065ae49c79f2dfcfa26e53e1fc9ad5421cdb60be8ab72b

Observation fb1b47f1-2cd9-4ef9-8f93-15d57347892a · outbound

This paper cites Transfusion: Pre- dict the next token and diffuse images with one multi- modal model.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Transfusion: Pre- dict the next token and diffuse images with one multi- modal model

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:35.139936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:32.668921Z digest=sha256:96d962d81392fa226572de219b0c7da33115b449bbd558d413650aa4520cce6d

Observation 98c75145-325c-4597-b82a-c9d1b8c10234 · outbound

This paper cites Uni3d: Exploring uni- fied 3d representation at scale.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale Uni3d: Exploring uni- fied 3d representation at scale

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:29:34.832315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:32.831965Z digest=sha256:2e2cd556f4aca8fa26a6341de3fcfeca69fa2ad04587c022f724fa9eeff03d43

Observation 691352d1-c632-4cef-8c1a-a8e3adc332f2 · outbound

This paper cites PartSLIP++: Enhancing Low-Shot 3D Part Segmentation via Multi-View Instance Segmentation and Maximum Likelihood Estimation.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale PartSLIP++: Enhancing Low-Shot 3D Part Segmentation via Multi-View Instance Segmentation and Maximum Likelihood Estimation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:32.980977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:29:32.980977Z digest=sha256:5abf3743ddc311e3e5ac7e47bea4ca3ebf3ddbfbc4ec6a118f2913359cc076b5

Observation 02869e3b-8dfb-4bfb-8f96-8e25208f20f0 · outbound

This paper cites embedded inside.

DexVLG: Dexterous Vision-Language-Grasp Model at Scale embedded inside

Reference 73

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:29:34.596928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T20:29:33.097470Z digest=sha256:43c3b78c2ad5a659176bc10e0385544e4ee62fe5a62b2223146d274bd5b6735c

Pith citing papers

Observation 8a4016a4-f65d-497a-8bd7-098b8b8eb7f8 · inbound

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge cites this paper.

DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:42:41.415484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T15:42:41.363422Z digest=sha256:45fbfaf1b30951e6299378c833fb944b51f0d8a34c4facf5fb003fcf75eb4984

Observation 9c96aaab-977f-4a81-9243-af342b15fcf0 · inbound

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos cites this paper.

Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:33:39.445246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:33:39.445246Z digest=sha256:6c555dffe73c81d95a2e1a83315e5f3145b36b78e8b0e7b6f3294a6740b5543d

Observation 71f0ccd7-34d9-47a4-8cd1-8c130286bba1 · inbound

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands cites this paper.

Learning Geometry-Aware Nonprehensile Pushing and Pulling with Dexterous Hands DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:52:38.724758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T13:52:35.939944Z digest=sha256:40814bcf1976d128d6ee8db768f5434cb424a5531d5ac06460f5cd34dd5fc24b

Observation efb12478-110b-431a-ab25-204dd742e19b · inbound

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models cites this paper.

AugVLA-3D: Depth-Driven Feature Augmentation for Vision-Language-Action Models DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:02:24.955096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T06:01:30.803128Z digest=sha256:e837b759cbc57833d558540d3493c4645541073c5f6090e7dbf588c5c0b8a665

Observation 47baa568-4a1b-4428-bc15-319925288dd0 · inbound

BiDexGrasp: Coordinated Bimanual Dexterous Grasps across Object Geometries and Sizes cites this paper.

BiDexGrasp: Coordinated Bimanual Dexterous Grasps across Object Geometries and Sizes DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:50:56.957350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:49:49.433772Z digest=sha256:b0d469163c02f53606d5794f0e3801f9ea810e0f463b24a80f8b1c0d53a2649e

Observation 5f05e673-cf52-480f-bb0b-9bc2bc9af562 · inbound

BLaDA: Bridging Language to Functional Dexterous Actions within 3DGS Fields cites this paper.

BLaDA: Bridging Language to Functional Dexterous Actions within 3DGS Fields DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:51:25.793833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:25:09.012355Z digest=sha256:f0e7bf1d7269003caf22ef1faa9b8a13a5a00a47d6788e0906a0923a00f2c327

Observation a7a90edd-ed5b-40ad-bdd1-bc8a9ed4a4b1 · inbound

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation cites this paper.

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T06:04:34.677454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-08T05:55:32.387354Z digest=sha256:fd05f6af11f6ac9b900a8befffa08a6b9d70a11e8d30a6a8e831ee021dcf4cb3

Observation 798058ca-a00a-4edb-b787-4ba0a6250bdb · inbound

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation cites this paper.

WristMimic: Full-Body Humanoid Control with Wrist-Guided Manipulation DexVLG: Dexterous Vision-Language-Grasp Model at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T16:02:59.358284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:02:59.358284Z digest=sha256:3467db4d07c151d103a2251db4d055997035140954b946c0d0f8c32d43653bad