Pith. sign in

Paper Citation Record · LEDGER

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection

As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2502.09829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.09829 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:24:08.254315Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy46
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 26c6ca63-413b-481a-93ed-bf280aed8ed2 · outbound

This paper cites Sim-to-real transfer for vision-and-language navigation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Sim-to-real transfer for vision-and-language navigation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.084157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.005910Z digest=sha256:a239958bc826ba629a2e4dfadd7fb9a82fc7a7c79f9f96c8652a08cc7fc11d98

Observation ed285a56-c073-4fc5-92c2-1360f2f1e3d6 · outbound

This paper cites Con- trast sets for evaluating language-guided robot policies.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Con- trast sets for evaluating language-guided robot policies

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.069010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.011688Z digest=sha256:b1c12ab2fc557ec17f77619f94dafe3e9f16d9ce03d56a16a61fc84962a2f80b

Observation a94521aa-8179-4cf4-a5a7-c1e53b0b35b7 · outbound

This paper cites Remembr: Building and reasoning over long-horizon spatio-temporal memory for robot navigation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Remembr: Building and reasoning over long-horizon spatio-temporal memory for robot navigation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.053902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.016879Z digest=sha256:440a5bce7b1a4c75f26f070cb0d61395c35f6aa0583cf906914b0a23c813b173

Observation b6c14229-43af-4668-9123-170bb3a08b9d · outbound

This paper cites Surrogate assisted generation of human-robot interaction scenarios.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Surrogate assisted generation of human-robot interaction scenarios

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.039893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.022672Z digest=sha256:00979fdbd4523cd46f0328ab87ddb40c15cd6682938784ca2811828bae76fcc2

Observation cc92b23e-6d74-4a9e-a3e9-948afe2cdf3c · outbound

This paper cites Mixture density networks.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Mixture density networks

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.025105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.027904Z digest=sha256:1c8ae9c2db880396c77f9236e9a120f19c697b24bd0f32ae0f4ea98f41c8329d

Observation db870277-d89b-4a27-89cc-73048e4422ed · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T20:24:08.032884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:24:08.032884Z digest=sha256:bd331e7b20323bb24cff2233b6480d410e7c9adb1338f76f05da097fc69872e8

Observation f22198aa-5c07-48af-a897-8d3fffdc7860 · outbound

This paper cites A Tutorial on Bayesian Optimization of Expensive Cost Functions, with Application to Active User Modeling and Hierarchical Reinforcement Learning.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection A Tutorial on Bayesian Optimization of Expensive Cost Functions, with Application to Active User Modeling and Hierarchical Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T20:24:08.038303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:24:08.038303Z digest=sha256:18f9033d53d7cc41207a8527403b1abbb3da40f146daec906c14f07e734fa724

Observation a89ec4cd-1752-49f3-9edd-5370b68f058a · outbound

This paper cites A survey on evaluation of large language models.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection A survey on evaluation of large language models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:09.010063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.043968Z digest=sha256:35561d4ff8263259f795c78edc39038ee1f458a75d8ad44175cf0ca6fe79050c

Observation 7af56f50-e073-4895-9362-7eaeeb9e849c · outbound

This paper cites Learning surrogate models for simulation-based opti- mization.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Learning surrogate models for simulation-based opti- mization

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.995807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.048745Z digest=sha256:18d575c052734a42b72931c0c3b5c85e45ec79d8bcd7b400fbec8568e6f91ca5

Observation 2559ab78-111d-49eb-900b-57a59fde8450 · outbound

This paper cites RoboTHOR: An Open Simulation-to-Real Embodied AI Platform.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection RoboTHOR: An Open Simulation-to-Real Embodied AI Platform

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.982018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.053466Z digest=sha256:668b4c36bc38bdf3353f872838d0c1499bfccf52ee44a5121af6a656893a6142

Observation 7cf5f8fa-e923-4c70-9727-6bfae3fff535 · outbound

This paper cites Efficient benchmarking of hyper- parameter optimizers via surrogates.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Efficient benchmarking of hyper- parameter optimizers via surrogates

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.964796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.058304Z digest=sha256:e4f49ea84d2d38d14e9d830f4e2432c69d44eca96df9759f059a00c184d584b5

Observation 55e5f08b-1cff-4d78-a5ab-059acbd87550 · outbound

This paper cites Dropout as a bayesian approximation: Representing model uncertainty in deep learning.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Dropout as a bayesian approximation: Representing model uncertainty in deep learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.950854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.063294Z digest=sha256:1a83e9d78af78a568c8a2a9873dcf442d3d4c8b28d4c894eb0b6eb6b2f81f065

Observation fadbd227-7ddb-47a9-b67b-8ce59014fc0e · outbound

This paper cites Efficient Data Collection for Robotic Manipulation via Compositional Generalization.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Efficient Data Collection for Robotic Manipulation via Compositional Generalization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.936588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.067842Z digest=sha256:2d2a8e46393d87152c58b75376ea45be4835a32ac4f887f0e0ba235af550995e

Observation 663b53e1-19c5-4ca3-abaf-0fcec1a5f52b · outbound

This paper cites Liu, Phoebe Mul- caire, Qiang Ning, Sameer Singh, Noah A.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Liu, Phoebe Mul- caire, Qiang Ning, Sameer Singh, Noah A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.922623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.072828Z digest=sha256:b7892ab22e8acc0102b148763a2da466c7d1944ac6018ea9904c45a21b1d39ad

Observation f9ebe5db-34ae-4c54-8d89-1eb480d460c8 · outbound

This paper cites Navigating to objects in the real world.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Navigating to objects in the real world

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.908659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.077336Z digest=sha256:9c588bc83b7d1e60476a481a591605107224d2feab74350d936af58c8c93edd3

Observation 7b9e03f1-767c-4715-9a1e-7908ecd69939 · outbound

This paper cites Minillm: Knowledge distillation of large language mod- els.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Minillm: Knowledge distillation of large language mod- els

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.894585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.081928Z digest=sha256:652f48020f1d84fdae6dffcafa8c0749f51531140652d30d9f23a89a60ef7db8

Observation 40d46ed5-1ef9-438f-b1ce-3ad86553ddaa · outbound

This paper cites World models.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection World models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.880292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.086506Z digest=sha256:7429e4b32d93fa8692bdb3a34d5bbfc2a46c41afbcfd1714852fdc7ffa48948a

Observation ae60c01c-38bb-4eab-bda8-933256a9e198 · outbound

This paper cites Benchmarking neural network robustness to common corruptions and perturbations.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Benchmarking neural network robustness to common corruptions and perturbations

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.866591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.091677Z digest=sha256:e9d09b7f52c2253b94a2b9f84c7d28b2ad2dd51ea88806e8d5514707643c9f04

Observation 3c0329c7-1abc-44ab-923c-bd8c08fd8085 · outbound

This paper cites Pretrained transformers improve out-of-distribution robustness.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Pretrained transformers improve out-of-distribution robustness

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.852914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.096785Z digest=sha256:f6e26bd3fa3ccfa641d2cb9c87c3b064b8931bdfc73620ec285e2d2ac92715d4

Observation c923236e-2c6a-4d33-9749-448d587481be · outbound

This paper cites Bayesian Active Learning for Classification and Preference Learning.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Bayesian Active Learning for Classification and Preference Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T20:24:08.101849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:24:08.101849Z digest=sha256:fcfc6fe8d611bb5abe8c704c4f01c9c617c7a9952db8551de770657fb342bada

Observation 60d4b5d4-e7a8-4930-b747-99cbbc76109a · outbound

This paper cites Deploying and Evaluating LLMs to Program Service Mobile Robots.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Deploying and Evaluating LLMs to Program Service Mobile Robots

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.838641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.107159Z digest=sha256:07293f45d99f212e5c56a9a4e68d103a7ca00160a20c63d0e31de95adb0f67b5

Observation fd54f599-94bd-4af1-a767-60d7402d995f · outbound

This paper cites Sim2Real Predictivity: Does Evaluation in Simulation Predict Real- World Performance? IEEE Robotics and Automation Letters (RA-L), 2020.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Sim2Real Predictivity: Does Evaluation in Simulation Predict Real- World Performance? IEEE Robotics and Automation Letters (RA-L), 2020

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.823846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.111866Z digest=sha256:7b27e4b1f803b80fdcb1381bc8376bb23df2bcc70dbd359057b604dfd5550c9d

Observation 0220736c-de73-4e23-a708-e1c0eed7aebe · outbound

This paper cites Openvla: An open-source vision-language-action model.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Openvla: An open-source vision-language-action model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.808906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.116672Z digest=sha256:9338f8d8dc91a38e09ab4d9fb003362372fbe5e8b3f0536f3fdf241cdd2f67f4

Observation 7254eecc-6edb-4afd-b687-2409a6c5afb6 · outbound

This paper cites Active testing: Sample-efficient model eval- uation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Active testing: Sample-efficient model eval- uation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.794038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.121494Z digest=sha256:6cfcac3dc90efb7a82430fd24b9e0d01d0cb8cc08a0b0b010ff70399526639e4

Observation 345be716-df4f-4671-a31c-23085459e484 · outbound

This paper cites Robot learn- ing as an empirical science: Best practices for policy evaluation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Robot learn- ing as an empirical science: Best practices for policy evaluation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.778709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.126153Z digest=sha256:39ba734d70ffc8e53c713c976f03eacc36c1c6cbb7371f86c068a90be63caf16

Observation f273a139-e0bd-4c76-be64-b1bd2191dfd9 · outbound

This paper cites Dropout injection at test time for post hoc uncertainty quantification in neural networks.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Dropout injection at test time for post hoc uncertainty quantification in neural networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.763886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.132137Z digest=sha256:1123085bcefa69a0a6df036885a8cccd79825b5a9bede1e787414dd4ae5054f9

Observation b966978c-e2d1-485a-9ba5-56afc1409e74 · outbound

This paper cites Cost-aware Bayesian Optimization.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Cost-aware Bayesian Optimization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T20:24:08.136868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:24:08.136868Z digest=sha256:8fed12a4156283a85c6eb8f4478ec437dd0589c801d358c2b8f203dbe83512a6

Observation f9782aab-e785-462e-9e64-ba8d3e0c759b · outbound

This paper cites Evaluating real-world robot manipulation policies in sim- ulation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Evaluating real-world robot manipulation policies in sim- ulation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.749435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.141947Z digest=sha256:b4484a040ddd4754b752436a823a178fdb1bc64aa94c19399e7a0187c00089d7

Observation e6f2b0d9-07de-42b5-b5a7-47db44b4efdd · outbound

This paper cites Ham- ster: Hierarchical action models for open-world robot manipulation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Ham- ster: Hierarchical action models for open-world robot manipulation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.735234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.146913Z digest=sha256:b0d5a886b4c454288a5e618b5f948e5290f55c73cb120d179cb330885c9f118f

Observation b9cab6f8-e7db-4cfb-96e6-2392c58d760a · outbound

This paper cites Holistic evaluation of language models.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Holistic evaluation of language models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.720648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.151565Z digest=sha256:161d315425310ff005502d0e019c44e61b5f68a6a9d654051ebaefb836709f11

Observation e4d9510a-19ce-4c7c-b63d-7146c9900ee1 · outbound

This paper cites A general framework for uncertainty estimation in deep learning.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection A general framework for uncertainty estimation in deep learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.706028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.155969Z digest=sha256:6043fa5325c1c5ae204992f5027e0ec9d79c1473ddcecdc7c3970b62b8ff33f6

Observation ec8a53fe-460d-453b-9e82-7eda7d2072d0 · outbound

This paper cites Probabilistic matrix factorization.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Probabilistic matrix factorization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.691595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.160717Z digest=sha256:02c135bf12c080490eb72885f116182388a9c6ea8ad1f7cbcc13a31004729053

Observation 0510ec4a-e5a2-4c0a-8277-218380b6ea71 · outbound

This paper cites Differential assessment of black-box ai agents.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Differential assessment of black-box ai agents

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.677074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.165429Z digest=sha256:578d3be4f470ebca937260030b762bf70a07198d479a9f7424620294000291f8

Observation 8d361def-f9ec-4b91-bbb9-e2cf162c2127 · outbound

This paper cites Octo: An open-source generalist robot policy.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Octo: An open-source generalist robot policy

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.661880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.170063Z digest=sha256:7999ddb1735ba49f6c0d8f0e01ee4fa07f3e1adba7f89d131ba8ceae797ffa3e

Observation 485d3982-4785-4bc9-9df6-b9ca7e4476ac · outbound

This paper cites Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Investigating the Role of Instruction Variety and Task Difficulty in Robotic Manipulation Tasks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.646307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.174469Z digest=sha256:65e3c323f94da93e7c55383e72896c1fb89e885ad7002b39264601b9f94ab9f4

Observation 6485b935-0ae8-457b-b4bb-7a710568aac7 · outbound

This paper cites Cost-aware bayesian optimization via information directed sampling.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Cost-aware bayesian optimization via information directed sampling

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.630414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.179124Z digest=sha256:b4033e84468c5bb41f06c4b52f3db40935d57fb309e29156aa96703f4feb8063

Observation c4a3c133-140d-4825-a647-ecf73f7f4213 · outbound

This paper cites THE COLOS- SEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection THE COLOS- SEUM: A Benchmark for Evaluating Generalization for Robotic Manipulation

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.613142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.183537Z digest=sha256:52093d25fe2d918e0ec5971a277452893509d0d2d11f871648809bca15a59521

Observation c1c848a8-6bb5-4efd-bc03-dc74cf668a9a · outbound

This paper cites Building surrogate models based on detailed and approximate simulations.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Building surrogate models based on detailed and approximate simulations

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.596637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.187971Z digest=sha256:3b08272bcec40b135a412d03787b21bbc1891ea3b7498054fc8daa86b90438a1

Observation 5fcfeaf2-47ed-4286-8fa4-f33c7eec3645 · outbound

This paper cites Modern bayesian experimental design.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Modern bayesian experimental design

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.581857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.192597Z digest=sha256:c16688275ad263b59ded2ce1ec53eb659a6aef66131738c37407f21941cb6c73

Observation 7bb24d94-39b3-4b04-bf42-c0f64301c6d9 · outbound

This paper cites Do imagenet classifiers generalize to imagenet? In International Conference on Machine Learning (ICML), 2019.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Do imagenet classifiers generalize to imagenet? In International Conference on Machine Learning (ICML), 2019

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.567419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.197577Z digest=sha256:223b2455295a8cf284803964ad6ab3bdfbe46c04c6517e8228df5dc7a1a0d0a0

Observation 3baddab0-c4cb-423c-bb0a-c2b21ac47fdf · outbound

This paper cites Active risk estimation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Active risk estimation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.550901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.202262Z digest=sha256:97c67c9a26840b71dc8e6eff55b7f05380bfb9ce2e65c304a66de4acedc6ae33

Observation 35f8747c-1dc2-4f35-b18a-4273c8bf4a8f · outbound

This paper cites Vint: A foundation model for visual navigation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Vint: A foundation model for visual navigation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.535117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.206920Z digest=sha256:8ea24e3ab889556ae855087093dc38690963fb431f22a2e4e23a0f54b26953a1

Observation 6e72152f-c96f-4b56-89ea-b38137f396d2 · outbound

This paper cites Lm- nav: Robotic navigation with large pre-trained models of language, vision, and action.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Lm- nav: Robotic navigation with large pre-trained models of language, vision, and action

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.518129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.211508Z digest=sha256:2bfa17f88ae55c57e711202156c8aa5d61e3d17a5250ddc1af1c5844ed41fc98

Observation dc026d25-b7dd-4277-90d1-d17e9346d0c1 · outbound

This paper cites Taking the human out of the loop: A review of bayesian optimization.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Taking the human out of the loop: A review of bayesian optimization

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.502523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.216126Z digest=sha256:190431ef8e62a41fd0ee1b664d889afe8b200975ccca8906f526720b6124e5c2

Observation 53efac60-5367-47f7-b1ca-6fe513fef398 · outbound

This paper cites Targeted active learning for probabilistic models.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Targeted active learning for probabilistic models

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-07T20:24:08.321102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.220721Z digest=sha256:88fd516ea980b06add0b17e90fd29729452a44a3951c7ee500fe19c29bd66222

Observation fe7e84a9-ab7e-4e51-acb8-099f44cdbdb4 · outbound

This paper cites Discovering user-interpretable capabilities of black-box planning agents.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Discovering user-interpretable capabilities of black-box planning agents

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.486723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.225896Z digest=sha256:9838148e8a0e0e212327675622067fb115278cd10edbe9a5552de0d5c5bd6552

Observation 899f423e-4ef5-44af-82b2-1b9f420132ad · outbound

This paper cites Autonomous capability assessment of sequen- tial decision-making systems in stochastic settings.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Autonomous capability assessment of sequen- tial decision-making systems in stochastic settings

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.469447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.230274Z digest=sha256:2af12f5b50645b8b576e3be39e72a11fffe635506951d5e1abea72cc78041f71

Observation b8e12e8f-c18d-43fb-bbde-97d163b158a2 · outbound

This paper cites How Generalizable Is My Behavior Cloning Policy? A Statis- tical Approach to Trustworthy Performance Evaluation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection How Generalizable Is My Behavior Cloning Policy? A Statis- tical Approach to Trustworthy Performance Evaluation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.452435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.234744Z digest=sha256:fcb28ecce21d3106094bdde2098351ba4401b706ad1bf3198ada34a834bd6744

Observation e85d0b1b-4eb4-4f05-8848-d96ae0fd9673 · outbound

This paper cites CLINE: Contrastive Learning with Semantic Negative Examples for Natural Language Understanding.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection CLINE: Contrastive Learning with Semantic Negative Examples for Natural Language Understanding

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.436050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.239673Z digest=sha256:54b12edda312ae3124d6390f882eeab08234677c2f820edb4b71dadf50763d91

Observation b65b5839-f159-4f26-a717-7a1f269544a3 · outbound

This paper cites Decomposing the generalization gap in imitation learning for visual robotic manipulation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Decomposing the generalization gap in imitation learning for visual robotic manipulation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.419828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.244153Z digest=sha256:da70a261ba57d4e8881c7b2b748e848dc8ef2e8530c6cccac9b2bd724683818a

Observation 27c5d9c3-fc00-4aff-b267-6a8fa47ed7a5 · outbound

This paper cites Sample Efficient Model Evaluation.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Sample Efficient Model Evaluation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T20:24:08.248881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:24:08.248881Z digest=sha256:65c99bda588d8ad26e2506abd17a8c6e56b8ecb4885db0de32ae67b4538a9bab

Observation b321e13e-317e-4693-8189-37ef7d9bdfca · outbound

This paper cites Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning.

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T20:24:08.403255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T20:24:08.254315Z digest=sha256:68cb0161bd8f460a4f35211a59bc289cbf9872513e9f17118c02174bb5febd4c

Pith citing papers

No inbound Pith citation observations are available.