Pith. sign in

Paper Citation Record · LEDGER

ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2505.09698.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.09698 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:49:10.769599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:07:27.171003Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f2ad57d4-4681-4293-8b70-f9451405badc · inbound

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction cites this paper.

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:10.769599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:10.769599Z digest=sha256:de5c3f4df11408607b53edd1c87d6aaa8157c7a1752ca8362317cafa3566ab5e

Observation dae45dfb-be52-4deb-8e53-24b791872864 · inbound

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models cites this paper.

BOP-ASK: Object-Interaction Reasoning for Vision-Language Models ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:00:10.728601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T19:58:19.309634Z digest=sha256:91094088ea92d892a58e231fa4a009d00f9fb90141fe34878f7cb8de71e714db

Observation 4c1a6adf-4af3-469a-b683-32c14542a0ff · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:52.195334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:52.195334Z digest=sha256:8dcc33b75d618084e2f69383c752471377be828b9db8252ba24a6d78344cbd03

Observation 216fdecd-5e46-45ae-a198-190b700de118 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:43:38.818008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:8917cc83d4e1808de2bb9079c993b0ee66afcb9b3dd3cace490a5373df275d23

Observation dc307761-1811-4e52-954e-bd759c6ee3f8 · inbound

Colosseum V2: Benchmarking Generalization for Vision Language Action Models cites this paper.

Colosseum V2: Benchmarking Generalization for Vision Language Action Models ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:33:38.663024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T16:32:32.885345Z digest=sha256:5f647ec4068317947edae8c9d1fda457ef42457982f4eb4d95d8453c25018631

Observation e141a78a-3ed4-4097-8013-c5719af18fce · inbound

When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA cites this paper.

When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:07:27.172850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:26:47.393632Z digest=sha256:045270ca8882273871b90366b75e858e241c2fa948782c5cb09b8405235a67f9