Pith. sign in

Paper Citation Record · LEDGER

Behavioral Exploration: Learning to Explore via In-Context Adaptation

As of 7 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2507.09041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09041 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:15:48.064835Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact3
  • verified fuzzy8
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ff9dbc7f-779e-41fd-89b2-78b7f97a0a6c · outbound

This paper cites OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.139432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.139432Z digest=sha256:be9aed127e446908a3ce391e8d67ec6c753678e9b4ea978f6e033b0ff5baed3b

Observation bfb40ee5-e43c-4cbc-b38f-6a6faa1a6810 · outbound

This paper cites pick up the cloth.

Behavioral Exploration: Learning to Explore via In-Context Adaptation pick up the cloth

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.824409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:48.064835Z digest=sha256:1b600a30062642a3882b8ff81a115842d702278140861a2de764009561d40f37

Observation a33d863a-8050-4f76-8902-4a36bb23ce44 · outbound

This paper cites (2024) for goal locations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation (2024) for goal locations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.849914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.872987Z digest=sha256:7b769a4e17cedeb858b917cdb52d4892a2d48577dd945a6623433c3da58e81c3

Observation 66839269-e730-4751-a862-0c5b0834200c · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Behavioral Exploration: Learning to Explore via In-Context Adaptation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.647136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.647136Z digest=sha256:ce5fc023f48feea0e6e42bd852d327c2ac421a98b3de26cc3665eb64506f94a7

Observation 73b157a3-f552-4a24-9015-be6a70485ba0 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-1: Robotics Transformer for Real-World Control at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.830579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.830579Z digest=sha256:878eb2d2c2a705db8c0feb8747550908f2711d2f0e8b0fd3831581e1c9e4d090

Observation f3370e52-5772-4fed-a042-fde90e6d9f98 · outbound

This paper cites Exploration by Random Network Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Exploration by Random Network Distillation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.984443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.984443Z digest=sha256:655f0d962ce6566f7d73e157ece06257752117fd0e009f6b2f526dcf1b6f09a1

Observation 02260926-b3ea-45e5-bdf7-eac5ac645540 · outbound

This paper cites Contingency-Aware Exploration in Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Contingency-Aware Exploration in Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.135671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.135671Z digest=sha256:888f44bbf809c0eff6f6333ec44edc4817d495e4b28835c04dd857e6d4c0887e

Observation 62fa7601-e0de-4b5f-a996-8178d3dc2b6b · outbound

This paper cites RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RL$^2$: Fast Reinforcement Learning via Slow Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.645441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.645441Z digest=sha256:d803cf1c232c32730e093e2f6cde4f2a5681d987485f5c28f9b9a280aa06a12f

Observation b3281106-560a-4910-a84f-1bbf866c41c7 · outbound

This paper cites Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.040995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.040995Z digest=sha256:b5a2e6d30055c9701906f1d6c3cb8cab25f48a57726e15c8a995f6297e443c31

Observation 81de7d67-d06f-4447-99fe-9da5621dec8e · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.147549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.147549Z digest=sha256:67276d31199927fa54f8ae439fca7ac7abaaabb1cb914fc71838d9db5402f667

Observation 3a451d3c-9a7d-4613-a9f6-73a530d445ea · outbound

This paper cites In-Context Imitation Learning via Next-Token Prediction.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Imitation Learning via Next-Token Prediction

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.255368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.255368Z digest=sha256:be0fb26670c910c18a892273996f775218c117b4cd554e2c4bd9a7fc36a45ae8

Observation 6ecdf627-5f25-49ee-a91b-dc225feb31c8 · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.367192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.367192Z digest=sha256:a7a62c063b8ea5aadad007fc3a40225ea361073d36454d82c488320bd946422b

Observation 5cf61d41-82c4-49f2-a1c1-391bcedd557c · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.449930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.449930Z digest=sha256:c99284587feff5c95ab44a6e9a6d4fa30798c82db0d6582688b4d17da0b56efe

Observation e5aa5d05-03f7-44da-8450-2a974ed1891d · outbound

This paper cites RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.565027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.565027Z digest=sha256:4afd56a19ca35e6bd85eafeee80803b102496faba6fc064cb70c9f3b769f2053

Observation 146d1b07-f47b-41ed-b931-18e4a43e4af0 · outbound

This paper cites In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.651513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.651513Z digest=sha256:3de06321c4b22493f4dbf8a50be957c07c2ec5efe0e25c17f2b5ae24bf4509b7

Observation 795ba928-5177-4043-8b71-db7f03c7de79 · outbound

This paper cites Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Adaptive Exploration Strategies in Dynamic Environments Through Informed Policy Regularization

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.645862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:44.741297Z digest=sha256:102d2394e0ceaebf5e6093980b8b52d1e759bcd6dbb65432f462d2482f2636df

Observation 7470a545-9737-4273-805f-66e7c31b818c · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.853558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.853558Z digest=sha256:ff0c2533ba552c3f61e5593ca96eccb5803b34386881f49c57823a3809bccf0b

Observation 097a55a6-5e4c-4383-bec7-b4b810126b6f · outbound

This paper cites Can large language models explore in-context?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Can large language models explore in-context?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.938881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.938881Z digest=sha256:edd8d31205adbb36cbe27b9749c93873fafe164af0cd3e9968a7b9980997627b

Observation 0e7ef5bb-d240-44de-ad57-dd99ce411702 · outbound

This paper cites Reward-Conditioned Policies.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reward-Conditioned Policies

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.032209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.032209Z digest=sha256:060f34da2796f039b411a69e5a172c4e42e33d9a5daff27e09eba93920a139c0

Observation afbb608b-4a04-491f-af7c-92af6364fc93 · outbound

This paper cites In-context Reinforcement Learning with Algorithm Distillation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation In-context Reinforcement Learning with Algorithm Distillation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.169185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.169185Z digest=sha256:75cca567530102b5fa452fc06b4255fe47d56d235a40d010edf920c8b4313921

Observation bb7ae660-7817-463a-8d44-e2a7774f8a8a · outbound

This paper cites FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization.

Behavioral Exploration: Learning to Explore via In-Context Adaptation FOCAL: Efficient Fully-Offline Meta-Reinforcement Learning via Distance Metric Learning and Behavior Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.214018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.214018Z digest=sha256:a9faad9de75cbe2a6b8527f1da482a8953005b768f89cec5603c9d3bcafd507c

Observation 60e2fd1a-701e-4aba-b631-6bbda861eacf · outbound

This paper cites What Matters in Learning from Offline Human Demonstrations for Robot Manipulation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation What Matters in Learning from Offline Human Demonstrations for Robot Manipulation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.355493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.355493Z digest=sha256:9e76cd67e62e97479b8b3a0041a9cc44c16408d98f59d40d9fb158337923b3dc

Observation d485db85-a976-405f-82ee-fd8afe773c72 · outbound

This paper cites A Simple Neural Attentive Meta-Learner.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Simple Neural Attentive Meta-Learner

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.523780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.523780Z digest=sha256:4e431f5db7936e6dfae14731b7f4ae444ad671dc6cfb51147f25174ce1211c33

Observation 372ff9fe-6ebd-42de-9587-a664a68c9133 · outbound

This paper cites EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation EVOLvE: Evaluating and Optimizing LLMs For In-Context Exploration

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.636586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.636586Z digest=sha256:7b1e31a834fb6fad48418023fa713f1d252c73b4d85e31c3f72e96d68710abe5

Observation c3267b21-5bd8-48f9-92da-3b1fa1e45fc1 · outbound

This paper cites First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs.

Behavioral Exploration: Learning to Explore via In-Context Adaptation First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.786311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.786311Z digest=sha256:152018820ab2bd077bdce6d6ca566a29d3f185073af300de6df581ce07b3a3d3

Observation 31f13b59-f076-4a7f-a5c8-cc8841759093 · outbound

This paper cites Foundation Policies with Hilbert Representations.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Foundation Policies with Hilbert Representations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:45.975166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:45.975166Z digest=sha256:2e63c9149dc14cc8bf14b7b58e30c9fc0f2ae1a24704caf404ff78098cee26d3

Observation 7d2c65f9-3f61-4edb-9214-fe45be52d940 · outbound

This paper cites Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Vision-based multi-task manipulation for inexpen- sive robots using end-to-end learning from demonstration

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.875551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:46.166468Z digest=sha256:92aad02a2eb20f135374c07069ffb114f24b81053a698cb21f63da60fe3cd46d

Observation 90ae909b-c04a-42ea-9f25-23998aeb6cb4 · outbound

This paper cites Generalization to New Sequential Decision Making Tasks with In-Context Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Generalization to New Sequential Decision Making Tasks with In-Context Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.321234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.321234Z digest=sha256:aca8644145a5d3696593328c674e1b2120f0823bb495fdfba086344adc199625

Observation b162da14-636e-4fba-9f92-63a13628a46d · outbound

This paper cites Retrieval-Augmented Decision Transformer: External Memory for In-context RL.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Retrieval-Augmented Decision Transformer: External Memory for In-context RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.599472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.599472Z digest=sha256:813d57504ce0a7e6c45097bfdc633cc6a502034a9376cdc8d32c78dafe41cffd

Observation 0f89f866-955c-4bf4-ba38-85ca2722bb14 · outbound

This paper cites Parrot: Data-Driven Behavioral Priors for Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.657876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.657876Z digest=sha256:65ea5fd542d81b7b64f3a0771f0f216657723968a25d4246321b09c8227971e2

Observation 325a3a7e-8201-4ff1-971f-797eabf7a0e8 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Training Agents using Upside-Down Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.817188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.817188Z digest=sha256:d8cf3cdf84c79de059c6412a9f6f58e5973c883d6b4ce80cc2371c6ebe066a8d

Observation 9059c91a-0a62-442c-b61d-efe17c856aa9 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.931088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.931088Z digest=sha256:f29f0bc98bb2c2ea1970da0cdf3ee7e1420a5cc4fed6fdb91a1e6d03daa55723

Observation cddd209a-35af-4400-a6df-c07c7af80026 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Octo: An Open-Source Generalist Robot Policy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.007830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.007830Z digest=sha256:8ff4747f703ab95762262b88a880fd95da5824b7f9b76d6ede422e032de149b4

Observation b03a66d1-6948-4f7c-bda9-e273d4799f61 · outbound

This paper cites Learning to reinforcement learn.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to reinforcement learn

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.096413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.096413Z digest=sha256:5ff2c47fa3ab9f7a5c18da0c76e8b07a7d2463ce71a8507053504c1b535eed86

Observation 33a8a429-6049-4687-997f-bd57fc282f9c · outbound

This paper cites Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.138620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.138620Z digest=sha256:4613f338682d928f75c393365b8bd1849c79c2e7ba18672d603ec9e9a8ab040b

Observation 51883d58-9c0a-4517-8a96-8cde8c78d085 · outbound

This paper cites Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.433440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.237851Z digest=sha256:5fd95666bd1e613bf068713577e15b45ef43cbb4d306265d785010697d9048c9

Observation fb9ce7a3-292d-4538-8c65-9aa20660a83d · outbound

This paper cites Policy Expansion for Bridging Offline-to-Online Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Policy Expansion for Bridging Offline-to-Online Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.326956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.326956Z digest=sha256:50f605e7e09e09c1bef3a113bf39e94e1fe447648b8de0260111443824403115

Observation 37cdd30f-d2f3-4461-bc4f-23ce09c7276c · outbound

This paper cites MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration.

Behavioral Exploration: Learning to Explore via In-Context Adaptation MetaCURE: Meta Reinforcement Learning with Empowerment-Driven Exploration

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:15:48.248660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.417541Z digest=sha256:c4aca7ad45c237ebc3d24d848332f0f4cf86e20ea6cfa01b8b15295c39ebc143

Observation 61531bc1-3052-4de8-9495-2fb6aac61881 · outbound

This paper cites Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Deep imitation learning for complex manipulation tasks from virtual reality teleoper- ation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.866571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.503432Z digest=sha256:8124e7c526310bab06ce201d1cf5fe7d558561e2afbe043e4ad23ef03bc9ea6d

Observation c881c136-71dd-4d68-a29d-aef1647de061 · outbound

This paper cites Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.590155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.590155Z digest=sha256:4acec44287aa8eed97f2a0270b9f43273db995719391f49afe0c19b215af5d34

Observation 30bce4b7-83b7-4c2e-9f98-f9490b17a353 · outbound

This paper cites Autonomous Improvement of Instruction Following Skills via Foundation Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Autonomous Improvement of Instruction Following Skills via Foundation Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.644692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.644692Z digest=sha256:b25bd8e93e9f02d99da9ec4593879758a17f49f36fa6de3b6bb06cb8dbee65e4

Observation 2603762d-7459-40a6-8dc3-8993e2926b9a · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:47.730302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:47.730302Z digest=sha256:ca06e5c85ad09222fbf184214f679c71afd61fc543e0087a2f2a593efeadb992

Observation b4dfb3dd-112e-4bf5-8a52-9338548a61f1 · outbound

This paper cites The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The inner induction then immediately implies the outer induction step, that we reach a terminal state at episode k in ¯C t β, so that | ¯C t+1 β | = nβ − t

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.858265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.803849Z digest=sha256:a452cc9c466db3f64b4f5fc00237d375cd5766999d48f881931e3ae91aec2cf8

Observation 8279da63-2fac-4666-8a87-b6f5a0b2b396 · outbound

This paper cites For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector.

Behavioral Exploration: Learning to Explore via In-Context Adaptation For each task, we run with a horizon of 300 steps, and utilize the environment’s built-in success detector

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.832832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.983379Z digest=sha256:535f304d1d399dd172e87a529bd349980799eca389206129d4cb08465afb9133

Observation 0e00e78e-59f9-4e07-9a8b-919e2d910c44 · outbound

This paper cites As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen).

Behavioral Exploration: Learning to Explore via In-Context Adaptation As stated in the text, we evaluate based on the number of goals reached (for Antmaze) or tasks completed (for Kitchen)

Reference 750

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.841677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:47.929647Z digest=sha256:6bf71179a5174eff2442de2f9d4ac901c04c5f9e66f5b54e14e19bf5fbe66b84

Observation 06bbb158-007c-4972-8b65-d89688b22d1e · outbound

This paper cites K., Yu, T., Singh, A., Phielipp, M., and Finn, C.

Behavioral Exploration: Learning to Explore via In-Context Adaptation K., Yu, T., Singh, A., Phielipp, M., and Finn, C

Reference 2006

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:15:48.883292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:15:46.103102Z digest=sha256:eebab4b331b24bd5d9869ac94c3e99500fb20519980335c47c754b918f524b09

Observation 159f7b58-4c8a-4f70-9822-cd365b443697 · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:46.469662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:46.469662Z digest=sha256:c64072fa4144b69facd76b1b8d39bb30c3600f17c2bee0cdeabdf00ad9c3f325

Observation 16687d6b-fc2d-4674-b27c-a41e40e5c70b · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

Behavioral Exploration: Learning to Explore via In-Context Adaptation GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.438704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.438704Z digest=sha256:13f58ba6e7b10b11a608c9ec0044771b0f0931b4d4dee15e8a8d4d1dfb853c54

Observation 8dece67f-c87a-4143-9c04-e5050612bac0 · outbound

This paper cites Go-Explore: a New Approach for Hard-Exploration Problems.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Go-Explore: a New Approach for Hard-Exploration Problems

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.839731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.839731Z digest=sha256:ccbc96ee186bf3e2e266513789848cdedb012a3972046d11b2f74f7146eef202

Observation 5fec2223-9cf7-4aae-bf45-4ed6b797c76e · outbound

This paper cites From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data.

Behavioral Exploration: Learning to Explore via In-Context Adaptation From Play to Policy: Conditional Behavior Generation from Uncurated Robot Data

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.348787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.348787Z digest=sha256:32bae86b9f0ff17cb5c76617c0ed24f8ea50ce88cf1faf91c72947560551a47b

Observation 77a78522-f49a-4eb8-96e5-bd86295f3c90 · outbound

This paper cites RvS: What is Essential for Offline RL via Supervised Learning?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation RvS: What is Essential for Offline RL via Supervised Learning?

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.933445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.933445Z digest=sha256:06505585b0f8fa1859f535d2b7de9c7c2eb8c22084e3164b56c38b1367928dbe

Observation 4b748d70-c76a-49c1-86d5-ffaa173c1a62 · outbound

This paper cites Is Conditional Generative Modeling all you need for Decision-Making?.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Is Conditional Generative Modeling all you need for Decision-Making?

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.188468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.188468Z digest=sha256:eb1da32b759bc7a09f37088dc19e5b483476b58cbb9cc7084cbb1abd90f2ed69

Observation ef579c6d-ab4d-43b3-84fd-1ee3b1709d66 · outbound

This paper cites The Ingredients for Robotic Diffusion Transformers.

Behavioral Exploration: Learning to Explore via In-Context Adaptation The Ingredients for Robotic Diffusion Transformers

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:43.506860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:43.506860Z digest=sha256:afd24117cf0701a66f8104356c1207f9f4c30a3bd5381079a91a9bda2daac4f1

Observation eedd4742-33fe-498b-9d80-465effa9fd62 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Behavioral Exploration: Learning to Explore via In-Context Adaptation D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.885629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.885629Z digest=sha256:87178d3c3d0b5be2c3c7c506b300d62ae124618b24b6d147051d7a606759cb29

Observation 225f2833-5557-4fcc-893b-3bf91fd56b0f · outbound

This paper cites A Tutorial on Meta-Reinforcement Learning.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Tutorial on Meta-Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.311817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.311817Z digest=sha256:209c9f61b6abf6612b71baf86091e77e737e70b5913413e13f746bf010d858b1

Observation 3754c8bc-55be-473a-b402-3f05da810bf7 · outbound

This paper cites End to End Learning for Self-Driving Cars.

Behavioral Exploration: Learning to Explore via In-Context Adaptation End to End Learning for Self-Driving Cars

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.747298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.747298Z digest=sha256:bc7ea4b795096c7e722abe83803a4990a10efec52deec061bf0002fb68e6ad37

Observation 1029ab3b-869a-4882-a489-959be1ab0268 · outbound

This paper cites Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.529366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.529366Z digest=sha256:23149bd57b51bbb74150bf162fba350cf32be48d05bbf4a1a5a4a8a9575d2cec

Pith citing papers

No inbound Pith citation observations are available.