Pith. sign in

Paper Citation Record · LEDGER

Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2410.13816.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13816 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:58:11.380304Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3946050e-0448-4895-b3d3-9923c6d148d6 · inbound

Inference-Time Policy Steering through Human Interactions cites this paper.

Inference-Time Policy Steering through Human Interactions Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T12:58:11.380304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:58:11.380304Z digest=sha256:599167553729c3ce100b30ad231e0314f82d71f094593b30ad91b405eae37db9

Observation 82ba4deb-b497-40f3-8b1d-0234f7378ec3 · inbound

Steering Robots with Inference-Time Interactions cites this paper.

Steering Robots with Inference-Time Interactions Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T00:26:30.915268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:26:30.915268Z digest=sha256:cf9b2fd5081d1a75d4cc44cc5bd2c554ab0e12f3710c829c3c8f4182d0cb8fd7

Observation 431e4292-c8b8-49ad-8121-66766e9471ad · inbound

Steering Your Diffusion Policy with Latent Space Reinforcement Learning cites this paper.

Steering Your Diffusion Policy with Latent Space Reinforcement Learning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:55:46.465566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T21:55:46.183007Z digest=sha256:ad8ed541af0543c3491f142cb010d1f964b9d06ec41ebc687620defca3d0f50f

Observation 280bbe23-c9e2-4460-9622-240ff2a3765f · inbound

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution cites this paper.

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T23:06:33.836456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:06:33.836456Z digest=sha256:212165fd23aaedeee59a7775a40cbc211c324881463c1f1bfbc94f59df74a0e9

Observation f81a2e7d-54e7-48a4-a8d8-e0ffe599cf5c · inbound

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models cites this paper.

Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T10:30:23.621321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:30:23.621321Z digest=sha256:430dcfc5457fbc47bf8f6cbd99da3f2954e521fa3f2352c597e9430dc4d6775d

Observation 11f4e421-1fa6-43f2-9827-9479de4dc67c · inbound

EVE: A Generator-Verifier System for Generative Policies cites this paper.

EVE: A Generator-Verifier System for Generative Policies Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T14:10:36.223054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:10:36.223054Z digest=sha256:3c98499cc90bfdfb9f7c0d76a5d02015988f910e7f91437be52d1efead8aca49

Observation d6b58301-bdf1-45e7-ba8b-de75f3fcc548 · inbound

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector cites this paper.

You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:55:24.183504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T09:53:45.151347Z digest=sha256:844e25169e42f3a6d038d123e3c3a847e01f4df312b13e659c96528c6d93288a

Observation 5de3b609-1808-437b-a62d-46500b7d2574 · inbound

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning cites this paper.

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-13T16:10:12.689957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:10:12.689957Z digest=sha256:7f95e4c6965c2e1485a24980dd167d88ab0e0d6c1fbe1368fc9c70533948c2b6

Observation d27af200-99d9-4546-9045-cac5167c09d4 · inbound

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry cites this paper.

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 87

Resolution
unresolved
no resolver link, observed 2026-07-13T14:05:26.303000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:05:26.303000Z digest=sha256:e2d1ef432e68a3577070b18045ff17972b84de0f1cf8d11fd3671cdc5a7f6c14

Observation f39066e2-4d65-4b9f-aefb-86064780b19e · inbound

Action Images: End-to-End Policy Learning via Multiview Video Generation cites this paper.

Action Images: End-to-End Policy Learning via Multiview Video Generation Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:50:53.460917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T18:51:05.206602Z digest=sha256:38739231f8f048a89bf7479c94ff390150fcdb10c4d1eca7101464223e6db905

Observation 7dc30270-99c5-4322-b1d9-d2f294d4f994 · inbound

Sumo: Dynamic and Generalizable Whole-Body Loco-Manipulation cites this paper.

Sumo: Dynamic and Generalizable Whole-Body Loco-Manipulation Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:25:58.872304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T17:38:49.399967Z digest=sha256:3705d50475dc3d5f07dad62ca98004ef690b764df194aa2a0055e8ef1fd27afd

Observation 5dc265e5-2b5e-4b75-8fc7-209bda580483 · inbound

FASTER: Value-Guided Sampling for Fast RL cites this paper.

FASTER: Value-Guided Sampling for Fast RL Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:48:26.504357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T02:47:36.475845Z digest=sha256:c322f59df77032e35e554e5f0333c8ea822dfcefa4ec5620a9b7c141398c4c2a

Observation 6d5654a9-30d7-47e3-958c-cc74253ea683 · inbound

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training cites this paper.

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:46:11.874198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T08:07:47.793895Z digest=sha256:5d3b6d01869b6a7548ba0d41dd6da33157969d64a3f8794a9ad906ef1cd4c812

Observation 71ccceb0-4bcc-40fb-9997-df0e21f18a90 · inbound

Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations cites this paper.

Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:21:47.687009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T02:53:39.060764Z digest=sha256:95085b6bd574f79d9817b55b0d4657a4c4c8d0a8e313d5b34d9894158cf0437c

Observation 707986f2-bb99-456f-9771-85806412c125 · inbound

Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations cites this paper.

Agent-Centric Observation Adaptation for Robust Visual Control under Dynamic Perturbations Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:50:49.832076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T00:48:43.992238Z digest=sha256:8e33a5b5c4df316738a0c52df07f06d54bb01016fa878da42e935cbbb8afe2fa

Observation b2dd4fbf-ab64-4608-8ed9-e594cb534089 · inbound

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model cites this paper.

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:46:06.365027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T15:10:16.533927Z digest=sha256:a07e2769074f5be53c0882db9d0868b7dc4c30ab10c84be75aa3e327924cbc73

Observation 31736ac4-3980-4425-82c7-29c41eb16ec6 · inbound

Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation cites this paper.

Offline Policy Evaluation for Manipulation Policies via Discounted Liveness Formulation Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:17:06.358956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T02:14:57.505609Z digest=sha256:1ceb2c07a2748c4debb1316f09220b3d95cc1f1e9104c03c319bccf1cc4c44c7

Observation 47de585d-e6c3-4970-81bb-65a0411dbc49 · inbound

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies cites this paper.

BOKBO (Best of K Bad Options): Calibrated Abstention for VLA Policies Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:13:15.444767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T08:08:00.917021Z digest=sha256:fbb569d67dd632d9446828e2494c8e10850da0d645d2cac4fc4275a2f1e1d18c

Observation dc403619-571b-4905-a418-88591f34805c · inbound

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models cites this paper.

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T06:17:41.831432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T12:47:35.314486Z digest=sha256:e3a817c92e3b81ff56ad34620e2b5e1f6eea96f7a4aed08f156f632f7f72dc81

Observation abaf5e70-0ae9-4ffc-84d3-fedd7f610078 · inbound

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning cites this paper.

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:36.927107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T14:05:01.073951Z digest=sha256:9417a95e1cd9e2ad7801dc0fed8069b0ad5cecfac74884aaaf0eeccb1b4fcadd

Observation 656ee199-fd43-48f0-a3da-9094d7ce2ed9 · inbound

DF-ExpEnse: Diffusion Filtered Exploration for Sample Efficient Finetuning cites this paper.

DF-ExpEnse: Diffusion Filtered Exploration for Sample Efficient Finetuning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T01:29:22.899636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T20:19:15.766523Z digest=sha256:0a6809c19d1b8598daf3759001ffa7fae408aeece6d8c804a3f09b733c3db435

Observation 87600d76-9c70-45eb-b6d0-ed34b2eb15db · inbound

Robot Critics that Sweat the Small Stuff cites this paper.

Robot Critics that Sweat the Small Stuff Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:37.242543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T14:20:37.905355Z digest=sha256:ff8a7b362012c238eb8ed50b71174885ec890907d211546baeb9d4ff42cc57ae

Observation a5b6bf11-93d1-423a-b2bf-83fd5f440024 · inbound

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning cites this paper.

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:29:45.610956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T08:45:04.483217Z digest=sha256:76d325bcf042b473597964934fc2832c8cba7f9f7e09e50e292aced8b6fae089

Observation 07cfd9af-6379-472b-ad7f-1169317edb50 · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:44.522102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:7747a803400619625fa2c15661febbac1685f0602967fc09653597802742fa0a

Observation b412d2a6-f41b-43c3-80c0-dc84627e29ce · inbound

Inference-Time Robot Behavior Steering through Physically-Aware Reconfiguration of Task-Structure cites this paper.

Inference-Time Robot Behavior Steering through Physically-Aware Reconfiguration of Task-Structure Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:09:50.149613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T05:32:52.472573Z digest=sha256:2f0dd3be7a35f3f38331a0fc483aac8280f8d689da3eb26c3cc86ec435f9ecd7

Observation 503ae2ce-26a3-41dd-a1a3-2568040f645a · inbound

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models cites this paper.

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:04:21.089589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T06:02:15.781538Z digest=sha256:8e7c00277b991c85c1909d6ea2ef301dd2ea89587343c189b3c305287f63b213

Observation 3b62d44b-7c5f-4024-8512-23f5cb58520e · inbound

Adapting Generalist Robot Policies with Semantic Reinforcement Learning cites this paper.

Adapting Generalist Robot Policies with Semantic Reinforcement Learning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:45:42.717218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T05:09:29.625066Z digest=sha256:a95508c880a0590cbad091e7b8e2ae621c448dc3e634530deb5131527d0b01e8

Observation 0da31749-55ee-445a-a6b4-02101293d13a · inbound

DREAMSTEER: Latent World Models Can Steer VLA Policies During Deployment Without Any Finetuning cites this paper.

DREAMSTEER: Latent World Models Can Steer VLA Policies During Deployment Without Any Finetuning Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T06:30:52.991927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:30:52.991927Z digest=sha256:5d5e4df45eea874f90c326ffc6f2edfd55c9d748b3ed0459301dd3a1c8077d4c

Observation 24c08ec6-4883-4799-a82d-5fda3be8664c · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 86

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:53:50.136045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-07T12:47:29.552283Z digest=sha256:2be820915c3ac7aedf2b6f3c1832c6c2d67f506e6682a69a9489c43d23a6916f

Observation d644c3bd-b56f-4e03-93a7-2505c60d8a8c · inbound

LLM-as-a-Verifier: A General-Purpose Verification Framework cites this paper.

LLM-as-a-Verifier: A General-Purpose Verification Framework Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-11T07:02:51.850836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:02:51.850836Z digest=sha256:52feafcd817b0d2aad3774f7636b6ebaa17be1c837333a7fa3ba7398e33db3f4

Observation f6b1a3db-88dc-49e0-ba0f-d371be9e5e5e · inbound

Skills in Weights, Memory in Code: Hybrid Learning for Memory-Dependent Robot Manipulation cites this paper.

Skills in Weights, Memory in Code: Hybrid Learning for Memory-Dependent Robot Manipulation Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T17:49:08.632440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:49:08.632440Z digest=sha256:cd13224d5e11ad7d25a765f6b4834888ff89c36af2e490044f8ae4c4f92df5eb