Pith. sign in

Paper Citation Record · LEDGER

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

As of 16 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 30 inbound Pith citation observations for arXiv:2506.17811.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17811 v2

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:06:38.524296Z

measured 87 of 87 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:33:39.096472Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

57 of 57 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved54
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 4e059f11-7ec9-4262-8037-40809d52d008 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.720618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.720618Z digest=sha256:1c1db0758da5af0da3072a996d38bdf91ab3faa094205cad9deeb746114cb662

Observation 15066f09-9a54-4e0c-a180-b55bad10623a · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.820431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.820431Z digest=sha256:634d5f979470b377e270e5c85eeea3bd0dc8512a900da527750eec4ff4eaf29c

Observation 83488752-6157-4a42-85b0-71592f090891 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models PaLM-E: An Embodied Multimodal Language Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.917365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.917365Z digest=sha256:e2c0266f1c1328c586dea0fcb1b563b59c31a2e7212ae2363c1d8bef9588f83b

Observation c1ddb712-20ae-47fd-b766-afad2aa402dd · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models RT-1: Robotics Transformer for Real-World Control at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.945432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.945432Z digest=sha256:975e47cf738f1b52a964e442114dd73b17cf2e3459814216ee81b26a2af9910c

Observation c38bf6fe-0a3a-465e-b4e9-ef30531c2c0c · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.965394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.965394Z digest=sha256:3c90af6195ac94bca574254aacbd81c1296cde69ed30d0bf83d9df09c6b37a17

Observation c6df4f1c-18c7-4f44-93ce-4c05d9ee50cc · outbound

This paper cites DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.971689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.971689Z digest=sha256:43c595319f52a7b94a19e065b642a57b1283849f5b50cc9a60016c402a4671b3

Observation 91c077f5-fbb1-4816-a541-b54e26403883 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.977603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.977603Z digest=sha256:3ac200b70e0814d989d2d32eefdbf983d865062c7e06f4b6d809d5d1a789bcf6

Observation e7b01b71-4f62-40ff-a7b6-2c908f3ed1da · outbound

This paper cites Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.981890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.981890Z digest=sha256:7d71e461aa51fb3ddbd7d5078f1522befd93ddfa3e64ab7c8e991213e880434b

Observation 5173db46-84b5-4d06-8e7a-ca875a913f41 · outbound

This paper cites Real-Time Anomaly Detection and Reactive Planning with Large Language Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Real-Time Anomaly Detection and Reactive Planning with Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.986849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.986849Z digest=sha256:8adc56310934b31fa3be24f50f36b1f437eeb71f14c831a0b7c73aacbf947ae5

Observation 2bdb5384-3bac-4ea7-91d1-b8326729e01c · outbound

This paper cites Autonomous Improvement of Instruction Following Skills via Foundation Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Autonomous Improvement of Instruction Following Skills via Foundation Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:37.991892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:37.991892Z digest=sha256:dc9ac2a05b309a97aab4ba2c55f593ced5654f53faee60e00a4d12b8fd7020ab

Observation 3f30a8b9-0bfb-4ee3-b866-7f54b211250e · outbound

This paper cites an unresolved cited work.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:06:39.420120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.092611Z digest=sha256:2fda73c95ac1c2767c529f4ab37a9a42afaa4f2eee4e0750023ab1ee496268d8

Observation 253698f0-fb00-4a30-beaa-449d6f8d59a5 · outbound

This paper cites Re-Mix: Optimizing Data Mixtures for Large Scale Imitation Learning.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Re-Mix: Optimizing Data Mixtures for Large Scale Imitation Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.138469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.138469Z digest=sha256:59de3324892b2bfffbed5873c50103c2d35ce86b5918d01a5123b4d1ee816667

Observation 8a04da67-f2c0-431f-a736-e4372c90c757 · outbound

This paper cites an unresolved cited work.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:06:39.406138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.185687Z digest=sha256:4390a2266d55683ddf867ef9c06e941e245e6f023f61c1b4cbb7a9380dfe2a0b

Observation 0a58b227-3677-4a19-86b6-51e4a7e85b0c · outbound

This paper cites GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.206319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.206319Z digest=sha256:9ef9b57a9f05b769860d26e909693d262356b7b8a1b6bc88b3fd8c4d1a854dbd

Observation a58c5d9a-dbda-478b-9afd-3733512a29e4 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.210628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.210628Z digest=sha256:c519ac518b814b98a3954a692d4c5d775c5e373c47f53aec5158a4cd6a7bc246

Observation 38d57aa7-d733-4825-9016-d0a8f519a14a · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.214757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.214757Z digest=sha256:1f764d86601b5a8ee0ddcbc49ec757e7e7d1a98609d403d348877ffd96509f1c

Observation f90b4e2e-cd28-4276-86e4-075fe468485b · outbound

This paper cites Clark, S.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Clark, S

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:06:39.392652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.219134Z digest=sha256:b4c73104d27341bd62d3541dce060dc66bc143db636ca56fc351374935117173

Observation 6feb5e24-d370-4a0a-96a2-060a380a1766 · outbound

This paper cites CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.229182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.229182Z digest=sha256:64f265d2c52ca69e8c279dec3407c310bfcc1ae9753040398f499ad53070bf5a

Observation a6a72b8a-6b33-4917-b966-3427af5083b2 · outbound

This paper cites GRAPE: Generalizing Robot Policy via Preference Alignment.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.234157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.234157Z digest=sha256:9d8da4e3303b0fc60b927887063ab66cd920e8f134b434a244c133ac139408c8

Observation 1ad6c9ec-260e-4e88-b27e-e5c084e30ad0 · outbound

This paper cites SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.238907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.238907Z digest=sha256:dc8a88c7b4b1508ebe7a63ee76a8bbbf77e08de4769db0d6ad1454755e5c81c0

Observation d114ff00-4823-40d1-985a-fa979fc0fa12 · outbound

This paper cites an unresolved cited work.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.259743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.259743Z digest=sha256:d8d50ba0f0abfe31ce51c79c3ef9aa0708e091f1f5c24331304b2da98845a304

Observation b0674c40-77d7-4e80-9047-b2c0f26a802f · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.295278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.295278Z digest=sha256:00e8c13bf24cca6e54033c8b323836699fdbae52e67a9e35c4d4c2102a1ed1c0

Observation 2c39bab2-ea9c-41e8-acff-71b86cb73548 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.340900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.340900Z digest=sha256:53cb944a4d55803ad4ef2e5b505b223b0ab63c11d9187d4fba9ed0151e62357b

Observation a66afff2-0ada-4a60-8afc-cc57ebb4fa02 · outbound

This paper cites Archon: An Architecture Search Framework for Inference-Time Techniques.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Archon: An Architecture Search Framework for Inference-Time Techniques

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.373143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.373143Z digest=sha256:af0fb9b7ed418d4b92b031aeefdc1c129de36a64650198c0f3a73e66cc08ec11

Observation e2917ace-fb17-4223-82c4-71dc3e5cb266 · outbound

This paper cites Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.379078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.379078Z digest=sha256:730f66b0885b0884e63bf73eb71ffe722d899003256cf923d168ab9599643201

Observation 8c758267-f737-4aa0-b13d-33eac92f9687 · outbound

This paper cites The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models The Good, The Bad, and The Greedy: Evaluation of LLMs Should Not Ignore Non-Determinism

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.383966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.383966Z digest=sha256:21a71445b222d07b2249cafd5a87078a0cf7ca74659e6a11ee2cd04d05ba1a85

Observation 1c4e0469-4a93-4cc2-b050-1d0d756e64ad · outbound

This paper cites S*: Test Time Scaling for Code Generation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models S*: Test Time Scaling for Code Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.388648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.388648Z digest=sha256:cfd68ab90220eacf66eca3cc016a049126de3c754b254d1817376b697a022def

Observation 6938cea0-ed60-4892-90e3-3e03d74fda0f · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models AlphaMath Almost Zero: Process Supervision without Process

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.393156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.393156Z digest=sha256:4ebd39025777b1e4ea145c5c08005de0c41ae4f13aa1d2a0e24200debde96fe7

Observation 19bf0dd0-dd21-4d09-b833-3d0e400d2149 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Training Verifiers to Solve Math Word Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.397965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.397965Z digest=sha256:2bfe4747d3f193794486d25f6afcacb3d99ec5e8841dee81c2367288daa4629d

Observation f3642f3b-331d-4dbb-a3f4-9324325275fb · outbound

This paper cites How Do Large Language Monkeys Get Their Power (Laws)?.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models How Do Large Language Monkeys Get Their Power (Laws)?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.402486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.402486Z digest=sha256:d6b9b886dec9ce71c7c45b45cc45654da5172216078bb8129f8cda5484b7b115

Observation d0f22b34-486f-4f03-bba3-6eb13a59a4eb · outbound

This paper cites Neural Scaling Laws in Robotics.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Neural Scaling Laws in Robotics

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-15T19:06:38.882296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.407250Z digest=sha256:fa2286f4140ed7737cce185b43a47ac47dc1b4867eff2078e82315ae8540d7fc

Observation 9fc6d12e-f6cf-4ad7-aa15-8a4ab87c3080 · outbound

This paper cites Data Scaling Laws in Imitation Learning for Robotic Manipulation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Data Scaling Laws in Imitation Learning for Robotic Manipulation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.411781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.411781Z digest=sha256:f1afecf2f8212e0aff196e811ecdef624efdd15263397d898322c0727036bab3

Observation 03b07d9f-fe95-4199-9249-c0d9cefaa9a6 · outbound

This paper cites CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.416217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.416217Z digest=sha256:a1a2e1daa0cacb3262b807a3450664fa0a6f700f9fc5f8e95358ba67f1d4c757

Observation 69c0694f-c741-4fdc-98c6-6789447b173b · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Octo: An Open-Source Generalist Robot Policy

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.421626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.421626Z digest=sha256:ea0cc72acab0b439f16c99a494004f298f03ce22ab99c21ee9421e1abc968e38

Observation bb428566-5c6e-433d-9ef3-e2d51b6ab430 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.426202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.426202Z digest=sha256:5970a9c1cee8af6f09d488f13604461f086868b31e32c7383fbf0b22941fbff2

Observation 87cd5ef0-0eb8-4d50-9671-1aaff25bda4c · outbound

This paper cites Training language models to follow instructions with human feedback.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Training language models to follow instructions with human feedback

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.430602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.430602Z digest=sha256:c816ceb7e338dcbff721e08c1e5e21f7f34a0b9d12c12837b98c7665dc0fa22b

Observation d593b6f9-f771-40a1-a0c3-42734a85763e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.435136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.435136Z digest=sha256:4813527c8b4b8966eb709c7211ac1dbdb7e4b3a7b8fa718797c9f248c9a1c0cc

Observation beca4e98-53d6-4cfd-9def-6542eae23843 · outbound

This paper cites an unresolved cited work.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.439546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.439546Z digest=sha256:577e9372311bad462c7dd2a34bab68e4078368910ac67ff34a3f64948ca52e7b

Observation 285859cb-de5c-4421-9be6-28cf8b7069af · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.443789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.443789Z digest=sha256:b3b571e180481966399cc79d708c141c01215eb705a5f7fd5a69f6c322816504

Observation 36f6c4c1-bb92-4dc8-ad46-bf14e2487a9d · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Learning Transferable Visual Models From Natural Language Supervision

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.448239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.448239Z digest=sha256:18e7442ad9a7f45dbfb3812917f2e47cf6bed8fdbce14ff13946645a8811885b

Observation 0ba898ea-adac-46f1-9acc-c2a4af4ecb79 · outbound

This paper cites Evaluating Real-World Robot Manipulation Policies in Simulation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Evaluating Real-World Robot Manipulation Policies in Simulation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.452789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.452789Z digest=sha256:b2db7d1d0f99f38308913a7c5bae7b1bf3ac91460896f5cb2afb50c0c0b0b554

Observation 605e53ba-8823-48bc-92f6-e29625f49a38 · outbound

This paper cites Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.457661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.457661Z digest=sha256:1e235fc9e5b6b76d0027f7d6fd96686c85951ae537f6feea5cc467afb8f09892

Observation f1073a59-2748-43af-8613-ab36b3572abb · outbound

This paper cites SGLang: Efficient Execution of Structured Language Model Programs.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models SGLang: Efficient Execution of Structured Language Model Programs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.462357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.462357Z digest=sha256:8bd69cef0a7f7296dbd5c7d21d6033724a4119c2b0a463725d572e3c1688277b

Observation 382a5456-6d14-46cb-a901-946196f6c3da · outbound

This paper cites VIMA: General Robot Manipulation with Multimodal Prompts.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models VIMA: General Robot Manipulation with Multimodal Prompts

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.467452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.467452Z digest=sha256:0afd9ba6a83674700e5816277acaa6398b5e12b328b3a7577b22e9c6204579fd

Observation 3c70256e-bdb6-449a-aece-b9fae0ba27f3 · outbound

This paper cites an unresolved cited work.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:06:39.370955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.471692Z digest=sha256:727a0f42ace4651011e2007fd3c3c5c46923998c08d1bc1f95d2eff9e36a65a4

Observation 187590a6-9f47-48aa-8246-2f8484aa763d · outbound

This paper cites V aswani, N.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models V aswani, N

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:06:39.357391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T19:06:38.475810Z digest=sha256:c74f6bfcc3b2cde9181b091f0c840034a4615859cf914c4f397561d89ee7e57c

Observation 76d6634f-8d37-4278-8b20-cc132a863eb1 · outbound

This paper cites A System-Level View on Out-of-Distribution Data in Robotics.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models A System-Level View on Out-of-Distribution Data in Robotics

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.480379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.480379Z digest=sha256:5222271ea88af582c8f401e3cd6bc44edbecf09f65dfca1e20827ee92f6b3595

Observation c5a3e6d5-ee0f-42bc-9903-a7f9a22523fb · outbound

This paper cites NeRF-Aug: Data Augmentation for Robotics with Neural Radiance Fields.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models NeRF-Aug: Data Augmentation for Robotics with Neural Radiance Fields

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.484785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.484785Z digest=sha256:4eb89775a005b14e7ef97452e71740ad693c05e5ebbe285c953433e46b52fe4a

Observation 5c6cf54a-f027-4187-9659-7091f51b78ed · outbound

This paper cites Data Augmentation for Manipulation.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Data Augmentation for Manipulation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.489585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.489585Z digest=sha256:841298bbd621a6a101eda25a7a6373a5efc640d4d07e3fe3a3e80ddb5b80ffbb

Observation 494f23dc-975a-44cd-bac5-f8811053809a · outbound

This paper cites Code as Policies: Language Model Programs for Embodied Control.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Code as Policies: Language Model Programs for Embodied Control

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.494324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.494324Z digest=sha256:e80cd17e29c209ca539cd6d70628a91b90fcadc9364b79080922bc25d91e12c0

Observation 7b35c8ee-390c-4f4a-9fa7-3ffb0547b308 · outbound

This paper cites EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.499008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.499008Z digest=sha256:015d21e14a4e8b3273a5cffe2f60679edee787236420c8d0e60e368dd51b671c

Observation 91da8de5-3147-4eca-a9f4-a44fd29c289c · outbound

This paper cites Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.504584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.504584Z digest=sha256:6cc4a88ad2022e600a9d870e172e2cbe0c78f513e7fd09626e04ab8319dbbbd8

Observation 89a249db-1725-4a5c-9548-bd2031201532 · outbound

This paper cites From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.509271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.509271Z digest=sha256:e63fa1708110677574a0c6517bcfa250df7e17eb1d2a9506261eca45991f618e

Observation bcb67212-4abb-4fab-acc0-812546809658 · outbound

This paper cites Inference-Time Policy Steering through Human Interactions.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Inference-Time Policy Steering through Human Interactions

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.514099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.514099Z digest=sha256:9b872366a773125b806650bf1c57e1db4f42d86085b99081637a15ba7bca3f73

Observation e984e6f6-fe80-403c-aa64-3f8bbe667f61 · outbound

This paper cites CodeMonkeys: Scaling Test-Time Compute for Software Engineering.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models CodeMonkeys: Scaling Test-Time Compute for Software Engineering

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.519265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.519265Z digest=sha256:b138d4ff43e6e48683f3aec61adc504a9e8fab08bf6a501a31f56bdfdba5599e

Observation e40e821c-78c8-4712-bebc-2a8a37b30654 · outbound

This paper cites LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.524296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.524296Z digest=sha256:d0751cc902ac2efc5e02ff2c445f03ccb645267a47490778a9a4bedf815363ec

Observation 52c68ad5-b5b5-476b-9a8f-af757160255a · outbound

This paper cites Action-Free Reasoning for Policy Generalization.

RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models Action-Free Reasoning for Policy Generalization

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T19:06:38.223870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:06:38.223870Z digest=sha256:2243339cea9c4f59cdbeb10857be18815d5e7661ee6de4091777523a72ce26d3

Pith citing papers

Observation edfdb120-c5e4-469e-abf7-aaccd705f309 · inbound

A Survey on Vision-Language-Action Models for Embodied AI cites this paper.

A Survey on Vision-Language-Action Models for Embodied AI RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-24T01:25:54.534326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-24T01:25:10.150459Z digest=sha256:029e155e7b87c239c0f941cc46d181c0bef5be32c86f9adb2177a7a2e644b426

Observation a3653210-b34b-4cae-9497-566057a51829 · inbound

Verifier-free Test-Time Sampling for Vision-Language-Action Models cites this paper.

Verifier-free Test-Time Sampling for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T11:20:12.318041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:20:12.318041Z digest=sha256:23a7cb65f1344056251189980cc411f2139e12482746cbd027bb36f19318a2d3

Observation 5233924e-4ad9-46ed-9b19-e95e50ebb821 · inbound

EVE: A Generator-Verifier System for Generative Policies cites this paper.

EVE: A Generator-Verifier System for Generative Policies RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T14:10:35.494365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:10:35.494365Z digest=sha256:a3442d8a50b8ee07649ad90f02446f0269af64ea525676d5404b2598b9628ebf

Observation 5b96847f-426c-4cbc-bca9-6886648566a3 · inbound

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control cites this paper.

Steerable Vision-Language-Action Policies for Embodied Reasoning and Hierarchical Control RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:06:42.922905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T22:05:39.797848Z digest=sha256:18fbf2b0a245a2335f03886917810c22cbb674d56a134b7fe7312362067db18f

Observation 62c26e7b-1bfd-4e3e-b416-ec4a6fab46cc · inbound

When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering cites this paper.

When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:00:15.390421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T19:00:12.595469Z digest=sha256:4c1611e4a9ec6e524764fb34cdf91fa241160a0a61162086941da9a33dadde0b

Observation 6a5c4afd-13aa-458d-a997-ba4df1b255cc · inbound

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry cites this paper.

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-13T14:05:26.303000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:05:26.303000Z digest=sha256:ccf5f6b0f98dc35b3f67409b52d494988dc0c8dfd62659b665cbf5446598f4cd

Observation 5ed0ebed-4ab1-4060-8e6c-3a920776859f · inbound

A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model cites this paper.

A1: A Fully Transparent Open-Source, Adaptive and Efficient Truncated Vision-Language-Action Model RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:50:51.354393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T19:31:23.255452Z digest=sha256:70936eab96b75d142b2d94861fa98157bce20eaf138989d55a88fea5bbfad471

Observation e233ee14-b447-4839-b7fa-28a5b9b87253 · inbound

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models cites this paper.

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:06:19.436103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T05:56:15.244586Z digest=sha256:64f2ae554a41a84e7582a5f33ffdec695e9e7d2ce24b826f7567e6326b5f78f0

Observation 47fe3bba-3321-450d-8847-d759182c8b7c · inbound

FASTER: Value-Guided Sampling for Fast RL cites this paper.

FASTER: Value-Guided Sampling for Fast RL RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:48:26.489694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T02:47:36.475845Z digest=sha256:fc4759ec0e40000fc22a859f3613c1a4f71a9bf4a41440faea0806d74d782d64

Observation 7721425b-da35-442f-b07e-f7f3cd0731f7 · inbound

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model cites this paper.

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:46:06.411348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T15:10:16.533927Z digest=sha256:2a7101fd1e202cf6c1c20f9723dfdadaaf01c6afb53ba812a9ca6d0c4e42d078

Observation 58771551-4cd8-4a59-8dc8-ccb39886429b · inbound

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs cites this paper.

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:06:27.337192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T03:46:01.737920Z digest=sha256:b619203516adbfb8ca5949b7b9a57ef65bed03e4a812059b3b193f797e8ca416

Observation 4da659ed-a157-4656-ab83-5f7bda2d0144 · inbound

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs cites this paper.

Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:52:32.091527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T07:50:13.845621Z digest=sha256:324da3884e9f55c4a7644d981e2a4150b924c8b2e3f96578673bd2fd0fe46312

Observation 571dd059-587f-444d-8b48-6592f7ff21d9 · inbound

Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents cites this paper.

Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:19:28.884923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-14T21:04:05.443622Z digest=sha256:061984afb6fc514678399ba2e91aca2ac86e58dc9d4dec973913158fc2f1248a

Observation 25a87d92-38ca-45b2-85ca-2506e047f2b5 · inbound

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies cites this paper.

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.469458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-29T08:08:00.917021Z digest=sha256:231f3e71869e71f168510926add8fea2d920b9212e8b5d154b88828c7e1a8f9a

Observation 1e8edaba-78a9-4527-83e6-4658151a8536 · inbound

Position: Good Embodied Reward Models Need Bad Behavior Data cites this paper.

Position: Good Embodied Reward Models Need Bad Behavior Data RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-28T17:22:24.983591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-28T17:18:17.337336Z digest=sha256:12cb1c3f72dfb5a4b381801a5dc4805e8a49740083ecc8c7e5fd82869d4aee72

Observation d634f6ca-4601-4581-b9d6-b698e0d0fb38 · inbound

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA cites this paper.

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:07:23.740490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T19:57:25.010339Z digest=sha256:a24d577a94e8f4a6fc1861b11736c394b3c52c70ce66a89d8067b2d73821b65a

Observation 2a9d5541-b750-4a89-8d8b-960072f382df · inbound

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA cites this paper.

Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T12:13:46.079807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:13:46.079807Z digest=sha256:5837c0b7a35029561e89245d0fe42648db870498a72b8d7d544db28b48e7c7a7

Observation 249821c4-17aa-4189-b796-6c354f5e982a · inbound

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning cites this paper.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:37:25.336361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T19:36:57.231932Z digest=sha256:e17fd5c5d907d072b29c3412d5e93775a050cf8621a8b900213ce59b1a678bee

Observation 46650ddc-4c73-4c07-b549-9526b0621644 · inbound

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models cites this paper.

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-03T06:17:41.823491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T12:47:35.314486Z digest=sha256:3f454ac0915a4df4eec51f0994f5c217fe31c49e31ba849ad8f8a3e8eaa22890

Observation 1a19962c-56bb-46eb-b9d9-9650480934db · inbound

Improving Robotic Generalist Policies via Flow Reversal Steering cites this paper.

Improving Robotic Generalist Policies via Flow Reversal Steering RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:48:35.766856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T06:20:19.209180Z digest=sha256:840815d1b68cf64ad010ce6c5583d7107a87527e10ece0c5389bfc3840c57924

Observation f9191966-ddb5-46eb-ba49-0ba830711892 · inbound

DREAM-Chunk: Reactive Action Chunking with Latent World Model cites this paper.

DREAM-Chunk: Reactive Action Chunking with Latent World Model RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:09:15.344625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T21:25:03.953677Z digest=sha256:cd12ad4fc40b4a9d049acc2ea119a58644c234f90bf33a260650af4bcbb73603

Observation e1dc8254-71c4-45bc-8e92-7889ecbea1c7 · inbound

TEXEDO : Test Time Scaling for Controller-aware Language-conditioned Humanoid Motion Generation cites this paper.

TEXEDO : Test Time Scaling for Controller-aware Language-conditioned Humanoid Motion Generation RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:29:44.874350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T08:50:47.113217Z digest=sha256:cf48089001f0520811c15dd2362b17ff305ce42eed0266ea3abb4123724024ab

Observation 034156c2-f375-4932-95c7-767b4479c82c · inbound

FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation cites this paper.

FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:40:32.808464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-25T19:29:00.117285Z digest=sha256:e69d2458072bd74d970ddfb7bab1cfaa0c1d154271ec69b1d8ce9c4f7f11a102

Observation 24d9cc48-6a8e-483a-aba3-3455be10c0d1 · inbound

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents cites this paper.

Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:40:07.874571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-25T19:55:51.114244Z digest=sha256:782ffed2fe152763a20395078843878ba18836e8baf6a5db10a9ad14097b4c43

Observation cbf0cef8-129f-419b-99e3-a59a24afc55e · inbound

E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation cites this paper.

E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:59:51.743107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T04:48:28.868562Z digest=sha256:a9af269accf149328bdfdd4238e6336bd28528b005c8cb2105c9b04d0f18f19e

Observation a3dce2f3-a2d9-4f3d-a051-19829b317a00 · inbound

Sequential Planning via Anchored Robotic Keypoints cites this paper.

Sequential Planning via Anchored Robotic Keypoints RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T05:04:20.305003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T04:59:12.363425Z digest=sha256:2ee38b70556c2c953f7aedfddd3c86f5c63a5e0935187e29a92d68e3a8648cae

Observation 1ff96540-ca17-4162-bb44-c6ec42c22457 · inbound

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies cites this paper.

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-01T05:45:25.230116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-01T05:44:43.682828Z digest=sha256:fa086cb5139caf9f0d7d0490f5845acbb5bec68687b80e15724d1a495ce63aba

Observation 7dbbbd27-8f04-4a49-8abd-ca84d1638c6a · inbound

Addressing the Orchestration Gap in Generalist Robots via Physical Agency cites this paper.

Addressing the Orchestration Gap in Generalist Robots via Physical Agency RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T06:55:48.373261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:55:48.373261Z digest=sha256:304adc851b76bae6bb8061594338ff0da41010fd7030e9b91ce45cb0b47642d5

Observation 74547591-ffd8-4c5d-85a8-59a55fe33727 · inbound

Action Chunk Scheduling for Batched Robot Policy Serving cites this paper.

Action Chunk Scheduling for Batched Robot Policy Serving RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T00:44:25.466499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T00:44:25.466499Z digest=sha256:91ba9dec7d464e3c860cb20b743c77c025092c9ff3f36c26a2e06b994c11dfaf

Observation dd585a45-c7b8-4582-8fed-cfd5017f5521 · inbound

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models cites this paper.

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:39.096472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:39.096472Z digest=sha256:8049d5ca3dc0e7c83ee83a8b250c21a763c58e1faed976bb81341b8c523de213