Pith. sign in

Paper Citation Record · LEDGER

A Tutorial on Meta-Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2301.08028.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2301.08028 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:50:28.075344Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

48
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd31f41d-5841-42f9-bb41-4f9d1443c9dc · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey A Tutorial on Meta-Reinforcement Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:49.601815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:be84e0e3b9df7764ef5c47cca29f31cd9bcaf5c226207fd88f94f23914470db7

Observation 3a9e2440-3a00-405a-8262-17c0cc20c2eb · inbound

AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers cites this paper.

AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T18:52:13.713049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:52:13.713049Z digest=sha256:39101e139561051e08cfa6d5c7e655ae8da65bdf52ecd8587dd70959ccf7d1a6

Observation 0fd30c77-597f-4962-931f-db7762bf5ce3 · inbound

Mind the Gap: Towards Generalizable Autonomous Penetration Testing via Domain Randomization and Meta-Reinforcement Learning cites this paper.

Mind the Gap: Towards Generalizable Autonomous Penetration Testing via Domain Randomization and Meta-Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T21:51:38.292557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:51:38.292557Z digest=sha256:7fd1d571103b3d299fa6de7cc3cb0be459ab03549d5a663d45b3f35ce99a1137

Observation 2c7f9632-3f78-43d2-8622-bdf253c74ad2 · inbound

A Research Agenda for Usability and Generalisation in Reinforcement Learning cites this paper.

A Research Agenda for Usability and Generalisation in Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T05:58:45.689385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:58:45.689385Z digest=sha256:bf36a8fca6fccc48aa32e4a57e5a7ab38df7f638fe82a59c6866088ddb025725

Observation 827f82a1-f7ad-441d-a5ac-50b99948f016 · inbound

CausalCOMRL: Context-Based Offline Meta-Reinforcement Learning with Causal Representation cites this paper.

CausalCOMRL: Context-Based Offline Meta-Reinforcement Learning with Causal Representation A Tutorial on Meta-Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T17:09:25.527110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:09:25.527110Z digest=sha256:d7cfddeeca175e2d5b9e9c7f31be110e6c2e90ae9c4dd3d0b73c2a94939b314d

Observation ae91eb9c-5e3a-4e74-971c-2b444f76ddc9 · inbound

Develop AI Agents for System Engineering in Factorio cites this paper.

Develop AI Agents for System Engineering in Factorio A Tutorial on Meta-Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:20.446330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:09:20.446330Z digest=sha256:5ad19b9744a12cc27fcefe3992ea98e702579d01a73ef9d57397c25010a78170

Observation 85f5dfe9-ec56-472c-bd9a-99d729a56631 · inbound

Single-Agent Planning in a Multi-Agent System: A Unified Framework for Type-Based Planners cites this paper.

Single-Agent Planning in a Multi-Agent System: A Unified Framework for Type-Based Planners A Tutorial on Meta-Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T23:14:16.328951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:14:16.328951Z digest=sha256:e85853814856e901a1997cd243b76606061a56b6dc4ef55568cc7b98c8af2db7

Observation 69d9991e-1950-405c-bbc6-0ecfb25be90b · inbound

STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning cites this paper.

STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning A Tutorial on Meta-Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T19:10:00.113728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:10:00.113728Z digest=sha256:419010b0c5970bcdd9d5481f0c9ac57dd2c0b503bdb5973b3ab9aafd60b02246

Observation ce8f0fa8-a1ad-4dbf-8c38-a30ac0bcf42d · inbound

Meta-Thinking in LLMs via Multi-Agent Reinforcement Learning: A Survey cites this paper.

Meta-Thinking in LLMs via Multi-Agent Reinforcement Learning: A Survey A Tutorial on Meta-Reinforcement Learning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-16T11:50:28.075344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:50:28.075344Z digest=sha256:59d8fabd9578b2eeeffa91212b3a36bfe468eaa8dbab95dbcd620323fac7b8ab

Observation 88de1a95-ef77-40c5-a89c-ebfc37edfea0 · inbound

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments cites this paper.

Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments A Tutorial on Meta-Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T06:07:19.469725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T06:07:19.469725Z digest=sha256:6a2e0146f8cc9e8950cb1bd2fabea6ef6ecd4176d782a3dca3b423dfe210dbd6

Observation 682d6bb0-a33b-4e7c-b4fe-a1e51e5059ce · inbound

Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review cites this paper.

Combining Bayesian Inference and Reinforcement Learning for Agent Decision Making: A Review A Tutorial on Meta-Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T22:17:38.946121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:17:38.946121Z digest=sha256:30cebba8d854a449bf10024e48dc2fbe69862ebd85e58162961c57dceb1888b9

Observation f82a2d06-1f08-4906-94c1-092d16de2ef1 · inbound

Real-Time Verification of Embodied Reasoning for Generative Skill Acquisition cites this paper.

Real-Time Verification of Embodied Reasoning for Generative Skill Acquisition A Tutorial on Meta-Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:00:13.634142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:00:13.634142Z digest=sha256:ccd86e3eb8cf4c186d321bde49a146f125c8b030be76208971beb22a0e9b5b47

Observation 45f2e769-48f3-4c95-83cd-1b4cea426905 · inbound

Filtering Learning Histories Enhances In-Context Reinforcement Learning cites this paper.

Filtering Learning Histories Enhances In-Context Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:53.611965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:53.611965Z digest=sha256:f9d96d84f6b69947a93904494a2fed73ba4f8f5ba36d0e1b7407a73c134eb03b

Observation 14fff60b-933b-4eff-8483-c0f9dc237f61 · inbound

Fully Offline Reinforcement Learning cites this paper.

Fully Offline Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:15:57.942175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:15:57.942175Z digest=sha256:ec8d26022b3dda9114e4ea39eae1e73fe854a719c3ded6c2827b6fcfd99b09d8

Observation d3912036-18b6-4530-8977-ca777f3445fc · inbound

Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning cites this paper.

Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:47:01.535236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:47:01.535236Z digest=sha256:132c2cecc8538a84931fe24e96b64356b59e5aed31de4fb6810d75faf74e038c

Observation 60c842f6-b19c-4219-8770-a2d2879e6b58 · inbound

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks cites this paper.

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks A Tutorial on Meta-Reinforcement Learning

Reference 2002

Resolution
malformed identifier
no resolver link, observed 2026-08-07T05:52:47.134864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:47.134864Z digest=sha256:3f092ee7b885ff5f29c219f0304d6f796ccc0aea0f6312e8f453ea0fae6ccfdc

Observation 225f2833-5557-4fcc-893b-3bf91fd56b0f · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Tutorial on Meta-Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.311817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.311817Z digest=sha256:5a29502792a3cb0796eff435432f5380e9b477af3aa95fde5a5fa17ce68cdda2

Observation 59369bea-951d-4879-922e-ac071fdf0ea3 · inbound

How Should We Meta-Learn Reinforcement Learning Algorithms? cites this paper.

How Should We Meta-Learn Reinforcement Learning Algorithms? A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:48:45.339004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:48:45.339004Z digest=sha256:191b1d098bb83165f03a1d400bc0f630df3565dd1afbac7da77dc5918c3ce4ce

Observation 683289fe-91dc-4791-8999-9af641b436c9 · inbound

Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control cites this paper.

Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:56.168767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:59:56.168767Z digest=sha256:559b38e81fc1127b4f6992c68ef5378b63a04ac31fa25562bc8c34128bc402bd

Observation 828391e7-c6ee-418c-8edf-340adefe9786 · inbound

Adaptive Policy Backbone via Shared Network cites this paper.

Adaptive Policy Backbone via Shared Network A Tutorial on Meta-Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:49:59.772066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:49:59.772066Z digest=sha256:a283a5a30bfc0e4c887fa8edf03d61acf7a5ffc3d7745bc0f40f7c33ff9fd7d6

Observation c6e774a5-87ba-4673-8d53-e47ee57090eb · inbound

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism cites this paper.

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:48:51.163438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-17T01:46:16.923948Z digest=sha256:a8876e8ec05f5fb3ff9d44d41e310d1cf95bc28128ddf67387dbb0d02fa08795

Observation 98d02217-19bb-4280-9dea-a4aaf06a2ef1 · inbound

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions cites this paper.

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:31:14.988946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-16T20:28:46.739742Z digest=sha256:66cf9a2ecc859fa796ba9a47b62d80c424318b0e2fbf75c100664d28dec593f3

Observation 9abd82dd-6b71-4ffd-90e8-b502cdbc6eed · inbound

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning cites this paper.

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:10:13.258937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T14:07:10.387869Z digest=sha256:f373bdf94c064d5a71bd8801d6404e3c0c92197f487c13016ec09a90e8e17dcf

Observation 4238129d-7615-4b4d-9c33-b2ff00c67f38 · inbound

Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent cites this paper.

Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent A Tutorial on Meta-Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:46:35.716309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T20:46:15.275441Z digest=sha256:c491d5096d9358b91a19b960cb4744874a8931d7168619e42848b38c012d04c1

Observation 363f90f5-c360-4fe4-b15c-2e3d8c7c6886 · inbound

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management cites this paper.

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management A Tutorial on Meta-Reinforcement Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:44:01.443264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T18:42:50.962120Z digest=sha256:875a7cdec06d1de587cd56462bfc73ac2e24850064cd9e5c6ca17752b74a09de

Observation 53776305-ded8-4d2d-94b3-18c786f45477 · inbound

Certificate-Guided Evaluation of Reinforcement Learning Generalization cites this paper.

Certificate-Guided Evaluation of Reinforcement Learning Generalization A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T20:42:37.754644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T18:23:27.059117Z digest=sha256:c77350e5b65e490105378fe9fc5d996dbc3f8285c39e83476504594120dadd7e

Observation 6db16b8c-5030-40f9-b1c8-2c7c846815d2 · inbound

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning cites this paper.

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T13:42:27.758405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:42:27.758405Z digest=sha256:ff5add8d4c29a1570248670cf9d5d92d6c6dbb76976286dcbb0c170e890bc656

Observation 1426b008-29c7-463d-8e8c-353fc32af201 · inbound

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models cites this paper.

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.133896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T00:38:48.912290Z digest=sha256:29fe3274601a7ce6c03033be0ef215301825a4b917f792f107e5807d3c2e10ad

Observation e9a06953-df1a-4324-a09c-43cd8d594085 · inbound

Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks cites this paper.

Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks A Tutorial on Meta-Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T19:04:32.205336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T19:04:32.205336Z digest=sha256:3d279cf14c8c039738e487476ab31385f7d6d6c208b856813213046c698fed0e

Observation 03c40415-f36f-4610-b4ec-c8eb3f89f805 · inbound

ReBRAC-v2: The Return of the King cites this paper.

ReBRAC-v2: The Return of the King A Tutorial on Meta-Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T00:32:39.844330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:32:39.844330Z digest=sha256:600052872461417232625c8daf33a26d840fc19ea176587b5ec3ce81f279e7f7