Pith. sign in

Paper Citation Record · LEDGER

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

As of 6 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 8 inbound Pith citation observations for arXiv:2510.17801.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.17801 v2

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:00:36.314865Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:27:33.138727Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:47:41.129979Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d667513c-a550-433f-94da-44e012a285ef · outbound

This paper cites Three prompts are used to cover different robot types: single-arm, dual-arm, and mobile-manipulator robots, shown in Figure 13, 14 and 15.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Three prompts are used to cover different robot types: single-arm, dual-arm, and mobile-manipulator robots, shown in Figure 13, 14 and 15

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.225393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.225393Z digest=sha256:ef0a8c73a026678f1d726f2366c9e96ef0a1d2d38fd2da645087503bc2e2f52b

Observation 3dac44f3-67a7-4442-99ec-953a287a71c8 · outbound

This paper cites Two function lists are used: Manipulation (Figure 16) and Navigation (Figure 17), followed by a conversion prompt (Figure 18) referencing these lists.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Two function lists are used: Manipulation (Figure 16) and Navigation (Figure 17), followed by a conversion prompt (Figure 18) referencing these lists

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.230218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.230218Z digest=sha256:5718dc8e619e74d761490946b8376efe9e8ccee1ff40586b99de5aa0e862461a

Observation a4351abd-0bef-4cd4-9c6b-62aa1e7cf17c · outbound

This paper cites The prompt is presented in Figure 19.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain The prompt is presented in Figure 19

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.234198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.234198Z digest=sha256:c58596089f6c7911a3f94215f757a53f9a54356b0151c36b69b79cdef109851b

Observation 6b5c6ec6-2f8f-47e8-9532-d02ffb71f183 · outbound

This paper cites 21 RoboBench Figure 9: Frequency of each action name in RoboBench planning tasks.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain 21 RoboBench Figure 9: Frequency of each action name in RoboBench planning tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.239111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.239111Z digest=sha256:bffa350b52c100b024a625b992be8d8aba1543f093a7e78bee34569f98e246c3

Observation 9a1a724d-46d9-43d2-b77a-916f35a69c2b · outbound

This paper cites These structured annotations provide the foundation for evaluating compositional reasoning and downstream task performance.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain These structured annotations provide the foundation for evaluating compositional reasoning and downstream task performance

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.244189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.244189Z digest=sha256:4573d117619c8e8f3ef240073243f29d045a876fbbd63c5d55e77efa1a9ab261

Observation 1dca2ab3-8131-4bde-9d70-5eb51a1a57f7 · outbound

This paper cites This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.)

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.249445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.249445Z digest=sha256:1ffc2dfdc264d382924902fdebb4b619518c380d72e7b644ef99fc3e07d7239f

Observation 1d7fd7d6-90ab-4160-b542-5ac31a49375a · outbound

This paper cites an unresolved cited work.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.253874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.253874Z digest=sha256:a5685b7793635d51ec47597093c2365f24e6b54acaaa4e80f4d81b8f1d6993d5

Observation d5bd8d9b-1fef-49dd-ad50-5ea58423d80e · outbound

This paper cites task_summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task_summary

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.258872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.258872Z digest=sha256:1087f7a0458b80914a8f43c0b3d9a5c361cd2407a06674030301468a6993ebcd

Observation 9908cf5c-9e6e-4e55-ae4e-2997fd68188d · outbound

This paper cites Your first task is to accurately identify which hand is the left arm ([left]) and which hand is the right arm ([right]).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Your first task is to accurately identify which hand is the left arm ([left]) and which hand is the right arm ([right])

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.263613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.263613Z digest=sha256:9e902f9896d50d33ae9ef78fcd7247ab71eefa50c7ff863682c0784cf1b91e50

Observation b93b7e5a-9f5b-481f-a0b4-eb8f9ddfe2de · outbound

This paper cites This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.268231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.268231Z digest=sha256:46f2e9f2d4a2a603a225f2df02431573d68af22b37bce36c2aa1a9ca1cd13326

Observation 026a675b-a3e7-444b-ad19-2fef8d6f2f21 · outbound

This paper cites task_summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task_summary

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.273035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.273035Z digest=sha256:58c9a03d215ce7bdfa58867932077367fc8f99a748919ca1f749aea59054ee64

Observation 0282a178-a095-4a31-97f0-a815ff0ac538 · outbound

This paper cites an unresolved cited work.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.278201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.278201Z digest=sha256:63cf5ac6e4c40cab8f2954773fe64f4036f1592be95f1121a7d7b169ea66e8b1

Observation ef552541-521b-4611-a444-b2af21913cab · outbound

This paper cites Primary Tag.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Primary Tag

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.283081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.283081Z digest=sha256:4beee5c498ace196b3a897f7510ac4611909c3661c562a120e37c5ce1d809435

Observation 53e7a689-4731-452d-b6b3-9f001b84ae72 · outbound

This paper cites task summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task summary

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.287053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.287053Z digest=sha256:a987f1de009374c8b3be1b1668b6c50f7a10d98bb29d62a18eead268f3b6fd85

Observation 48902fc9-4523-415e-b4cf-0771f98c0040 · outbound

This paper cites plan step.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain plan step

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.292274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.292274Z digest=sha256:f984c2376cd9e3d8da665d37545fdd43c180238bdb5837f0a631caf6cd523d6d

Observation 22f04677-d00b-41e2-b6a9-f631545e506c · outbound

This paper cites reason" field explaining how the.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain reason" field explaining how the

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.297718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.297718Z digest=sha256:48749d94660140e581354e8ad2b6f4760e9f58334af6e4fcc02f90742aa537c9

Observation e2a64143-9eab-458c-a25c-9fa6d7adec80 · outbound

This paper cites Flexible matching in Standard Mode; strict in CSS Mode.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Flexible matching in Standard Mode; strict in CSS Mode

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.302289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.302289Z digest=sha256:e2fe4c0cd77ba7deef1eab135c961c2625b16a27920c1c916f3afc0f475047ee

Observation 89b42c28-34a1-4c92-a5d6-79f6a9984861 · outbound

This paper cites node_correctness.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain node_correctness

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.306227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.306227Z digest=sha256:560093010a30f99f2560727d3d4e47734677d24d7c5777e6d8349a6f65179957

Observation 47f5e79d-fadc-4038-8540-d4da43d07bb6 · outbound

This paper cites Award 1 iff the skills are exactly identical (strict match after normalization); otherwise award0.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Award 1 iff the skills are exactly identical (strict match after normalization); otherwise award0

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.311102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.311102Z digest=sha256:0b82a328c5ee618a03ea59aadba6e4021620f378a304814fa85609053e965132

Observation 31025de2-aefd-4c0f-97d6-1e70bf37a351 · outbound

This paper cites skill_usage_accuracy.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain skill_usage_accuracy

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.314865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.314865Z digest=sha256:e62239fa0fecd3303dd4ad4f01b41d8c83818a96b3da68987e80bbe5033439d6

Pith citing papers

Observation be51f4da-5231-4b67-ba5c-38f17e40056b · inbound

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics cites this paper.

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T16:27:33.138727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:27:33.138727Z digest=sha256:da2830730f536e224b8ad57a4443bd9c299af20ee982e459c52eae379dcb5300

Observation 3f4e115a-db87-41b1-a1fd-c8b5d6ec3795 · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:52.299425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:52.299425Z digest=sha256:e781b2f726bc2e84ae1f14120ab5ae26744185fb3d3e142d0a8c47823393610d

Observation cc83dbe7-2448-4b0a-81c0-00059d24958e · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:4d3db45af1feb46ae73d9acf67c9b71a11fae207566d3711e1f5640607ea86a4

Observation a8abd85b-21ab-4888-b00b-50d8bb0e32f2 · inbound

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration cites this paper.

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:17:23.965554Z digest=sha256:a349c87f31ff685488da9a5e5262eb991f4b161e0ead1df3cadb8e3b5312d67f

Observation 6a493a0e-a8dd-481a-a268-491f09f63228 · inbound

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs cites this paper.

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T11:02:07.122615Z digest=sha256:0cfc8f40b227bc3f525cd7d6d070deb83d97fae3064f5a0c3d01c82cf0e78ba7

Observation 89b38ef5-34e2-47ba-a8af-e4f612b55800 · inbound

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation cites this paper.

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T13:06:35.282837Z digest=sha256:ea5da20465a415e7511c5bb517496f1fcf7fbfcaab7155a1b1702413ca717101

Observation e5b6e00e-09eb-4701-ac5a-1b1d7c508c25 · inbound

An Exam for Active Observers cites this paper.

An Exam for Active Observers Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T21:12:05.853271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:12:05.853271Z digest=sha256:d97d4493cd94bdf78cb703952390505c11a81e77ebe44f708ede1118f8d39a7a

Observation 66e2f8f5-ce85-4015-be69-3d36cf187479 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:39.830975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:39.830975Z digest=sha256:9a69b256985fe44afc7df0878d97b89bb28b804e7fbdbf9a408e875cd2df177d