Pith. sign in

Paper Citation Record · LEDGER

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

As of 16 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 8 inbound Pith citation observations for arXiv:2510.17801.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.17801 v2

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:00:36.314865Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:27:33.138727Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:47:41.129979Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d667513c-a550-433f-94da-44e012a285ef · outbound

This paper cites Three prompts are used to cover different robot types: single-arm, dual-arm, and mobile-manipulator robots, shown in Figure 13, 14 and 15.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Three prompts are used to cover different robot types: single-arm, dual-arm, and mobile-manipulator robots, shown in Figure 13, 14 and 15

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.225393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.225393Z digest=sha256:3dd47043f072141f824b508308a5705e11b8d3f2187c221d447233fa28ad04d4

Observation 3dac44f3-67a7-4442-99ec-953a287a71c8 · outbound

This paper cites Two function lists are used: Manipulation (Figure 16) and Navigation (Figure 17), followed by a conversion prompt (Figure 18) referencing these lists.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Two function lists are used: Manipulation (Figure 16) and Navigation (Figure 17), followed by a conversion prompt (Figure 18) referencing these lists

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.230218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.230218Z digest=sha256:d840a2757f201f45b5161d6fca812516a4fb44b2510ee8057dbac8f998e207bc

Observation a4351abd-0bef-4cd4-9c6b-62aa1e7cf17c · outbound

This paper cites The prompt is presented in Figure 19.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain The prompt is presented in Figure 19

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.234198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.234198Z digest=sha256:befeecf4aec103013b4cbaf94c8121d889b1e1af7fba341788d69f8cd0b84cbc

Observation 6b5c6ec6-2f8f-47e8-9532-d02ffb71f183 · outbound

This paper cites 21 RoboBench Figure 9: Frequency of each action name in RoboBench planning tasks.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain 21 RoboBench Figure 9: Frequency of each action name in RoboBench planning tasks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.239111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.239111Z digest=sha256:1089b1bf334931b1a9386b67e2300b346971d8f546dd0290d980c74cb339f795

Observation 9a1a724d-46d9-43d2-b77a-916f35a69c2b · outbound

This paper cites These structured annotations provide the foundation for evaluating compositional reasoning and downstream task performance.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain These structured annotations provide the foundation for evaluating compositional reasoning and downstream task performance

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.244189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.244189Z digest=sha256:d8175f88dd0f32e9b3c1106b0bcdb7440bd75c35f3b18b96de96825b71225b21

Observation 1dca2ab3-8131-4bde-9d70-5eb51a1a57f7 · outbound

This paper cites This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.)

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.249445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.249445Z digest=sha256:35ef61019f6ca8edbb40e1881ed0d6fac39c6ca26089628dde1aff420a68e99e

Observation 1d7fd7d6-90ab-4160-b542-5ac31a49375a · outbound

This paper cites an unresolved cited work.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.253874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.253874Z digest=sha256:9ca56f8bb9e2c79f1d47422a3675e3b5515be0c41d383a8939bbf8a39a696116

Observation d5bd8d9b-1fef-49dd-ad50-5ea58423d80e · outbound

This paper cites task_summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task_summary

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.258872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.258872Z digest=sha256:ac009a61d1d26f211cd5fa67e10b2e96cf3117b9396bce80dce2ec3a89a63bae

Observation 9908cf5c-9e6e-4e55-ae4e-2997fd68188d · outbound

This paper cites Your first task is to accurately identify which hand is the left arm ([left]) and which hand is the right arm ([right]).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Your first task is to accurately identify which hand is the left arm ([left]) and which hand is the right arm ([right])

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.263613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.263613Z digest=sha256:08b1ec785ab3a9cf77bab61cf7e3382df7e6c49896bb676f03e6d29edae65207

Observation b93b7e5a-9f5b-481f-a0b4-eb8f9ddfe2de · outbound

This paper cites This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.).

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain This task could be a clear goal or a series of related activities (e.g., assembling furniture, repairing equipment, preparing food, etc.)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.268231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.268231Z digest=sha256:3cfe402299b3817cf603b5344ac6671b59481fc22ef65067c948068711c03bd1

Observation 026a675b-a3e7-444b-ad19-2fef8d6f2f21 · outbound

This paper cites task_summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task_summary

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.273035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.273035Z digest=sha256:158f38dc4120fe2263b84a234c9ecdb73ec9d0357f403133a0fd783a74699f55

Observation 0282a178-a095-4a31-97f0-a815ff0ac538 · outbound

This paper cites an unresolved cited work.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.278201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.278201Z digest=sha256:a2dbe3becfb786b8b17f289c8f18eb05d283ccc8d48ee96626dc6a9f72bf302e

Observation ef552541-521b-4611-a444-b2af21913cab · outbound

This paper cites Primary Tag.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Primary Tag

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.283081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.283081Z digest=sha256:3c5d1484bd87c831d8242f439fe19fb9870068a3fba8585a04a416a9a56af4e6

Observation 53e7a689-4731-452d-b6b3-9f001b84ae72 · outbound

This paper cites task summary.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain task summary

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.287053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.287053Z digest=sha256:b5b95be0cb34227a3b45ec676be615d803f07a217fbb25feb485f63bdac751de

Observation 48902fc9-4523-415e-b4cf-0771f98c0040 · outbound

This paper cites plan step.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain plan step

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.292274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.292274Z digest=sha256:ac47a699b3f50334e059e4966acdbecd6bf69983485fdda31b54d94dd24d74c1

Observation 22f04677-d00b-41e2-b6a9-f631545e506c · outbound

This paper cites reason" field explaining how the.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain reason" field explaining how the

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.297718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.297718Z digest=sha256:769980d2330f9975e411c857acfe24451ac415ebd7231abd19dc9cc31754c2f3

Observation e2a64143-9eab-458c-a25c-9fa6d7adec80 · outbound

This paper cites Flexible matching in Standard Mode; strict in CSS Mode.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Flexible matching in Standard Mode; strict in CSS Mode

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.302289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.302289Z digest=sha256:af369ce37821fc3c1278fe39ebf4295074308f4194dbdfa2fd9f7eb656fb31d2

Observation 89b42c28-34a1-4c92-a5d6-79f6a9984861 · outbound

This paper cites node_correctness.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain node_correctness

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.306227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.306227Z digest=sha256:ac915ff9a270e4c1b17ddba30ec37ebf0f5973fe70e642615708f4d2d6bac1e3

Observation 47f5e79d-fadc-4038-8540-d4da43d07bb6 · outbound

This paper cites Award 1 iff the skills are exactly identical (strict match after normalization); otherwise award0.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain Award 1 iff the skills are exactly identical (strict match after normalization); otherwise award0

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.311102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.311102Z digest=sha256:f0a410af502fedcf445e45a128284f076b34fba46e65d1d0f9474fac16c0b0bc

Observation 31025de2-aefd-4c0f-97d6-1e70bf37a351 · outbound

This paper cites skill_usage_accuracy.

Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain skill_usage_accuracy

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T09:00:36.314865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:00:36.314865Z digest=sha256:62cbc1765c607b4109d7c17e590a502114c3f98d590b07266d4633a6c07f95fe

Pith citing papers

Observation be51f4da-5231-4b67-ba5c-38f17e40056b · inbound

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics cites this paper.

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T16:27:33.138727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:27:33.138727Z digest=sha256:7e04e923418348acb3ecf43306d982eac472bb79503bf517e30545ec73f52dab

Observation 3f4e115a-db87-41b1-a1fd-c8b5d6ec3795 · inbound

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training cites this paper.

Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T13:29:52.299425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:29:52.299425Z digest=sha256:36c38a92124f7d5d051e44ac494580b676bd69c6cc3ee48a628aa6931ca46123

Observation cc83dbe7-2448-4b0a-81c0-00059d24958e · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:3059f1b116a150e8154f4ee75839daaadb186c963598e78c53729c259b1b6888

Observation a8abd85b-21ab-4888-b00b-50d8bb0e32f2 · inbound

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration cites this paper.

Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T23:17:23.965554Z digest=sha256:f582c09476e4f1a7c11e1a58a7f9a073711ebb49386f7596067939ee615cfb31

Observation 6a493a0e-a8dd-481a-a268-491f09f63228 · inbound

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs cites this paper.

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T11:02:07.122615Z digest=sha256:8a9eca9f6dccc676b055889ca9b86bdb8ca75098b5da64726f0b569aa89333fb

Observation 89b38ef5-34e2-47ba-a8af-e4f612b55800 · inbound

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation cites this paper.

A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:10.074313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T13:06:35.282837Z digest=sha256:009a0bdc047a908d12dd7e984adaee7c904617b51ac1422916d56478dda98974

Observation e5b6e00e-09eb-4701-ac5a-1b1d7c508c25 · inbound

An Exam for Active Observers cites this paper.

An Exam for Active Observers Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T21:12:05.853271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:12:05.853271Z digest=sha256:67552e91fd1814e6955705976b69f488f11645a16833994d84ed41845f1a1eec

Observation 66e2f8f5-ce85-4015-be69-3d36cf187479 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation Robobench: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models as Embodied Brain

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:39.830975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:39.830975Z digest=sha256:bbb8a52529947ecbed682f7a986a184f52ae049656f81034a325873b05d91816