Pith. sign in

Paper Citation Record · LEDGER

A Tutorial on Meta-Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2301.08028.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2301.08028 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:09:20.446330Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

48
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd31f41d-5841-42f9-bb41-4f9d1443c9dc · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey A Tutorial on Meta-Reinforcement Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:49.601815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:dd29ec39ac8f4f890010ea11d9ed1c109c56a4f5dd02751d0baeec27314730aa

Observation ae91eb9c-5e3a-4e74-971c-2b444f76ddc9 · inbound

Develop AI Agents for System Engineering in Factorio cites this paper.

Develop AI Agents for System Engineering in Factorio A Tutorial on Meta-Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T15:09:20.446330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:09:20.446330Z digest=sha256:adfdcf709faa72324a6fa8bff2dea7bbc7d7463feaf8d0e1ec37d70110c5a64f

Observation 85f5dfe9-ec56-472c-bd9a-99d729a56631 · inbound

Single-Agent Planning in a Multi-Agent System: A Unified Framework for Type-Based Planners cites this paper.

Single-Agent Planning in a Multi-Agent System: A Unified Framework for Type-Based Planners A Tutorial on Meta-Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T23:14:16.328951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:14:16.328951Z digest=sha256:8f60778a056db2b7ca77b4ba517a1420808225e0b25c552be4761358d066059e

Observation 69d9991e-1950-405c-bbc6-0ecfb25be90b · inbound

STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning cites this paper.

STMA: A Spatio-Temporal Memory Agent for Long-Horizon Embodied Task Planning A Tutorial on Meta-Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T19:10:00.113728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:10:00.113728Z digest=sha256:8ec428956b9706ca907a42fa5ff326cb08273c0b471d69dda8e6181cf600f6f3

Observation 45f2e769-48f3-4c95-83cd-1b4cea426905 · inbound

Filtering Learning Histories Enhances In-Context Reinforcement Learning cites this paper.

Filtering Learning Histories Enhances In-Context Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:53.611965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:53.611965Z digest=sha256:fa527888982b7614b81102f11c5a6ecd1369e4b236bdc1fbe7c07629425a4801

Observation 14fff60b-933b-4eff-8483-c0f9dc237f61 · inbound

Fully Offline Reinforcement Learning cites this paper.

Fully Offline Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:15:57.942175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:15:57.942175Z digest=sha256:ba84e251aa96ab08db4fa3cf51d2f2fa6077e062c6feb877d06444cf5bfb5bf0

Observation d3912036-18b6-4530-8977-ca777f3445fc · inbound

Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning cites this paper.

Unsupervised Meta-Testing with Conditional Neural Processes for Hybrid Meta-Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:47:01.535236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:47:01.535236Z digest=sha256:040b5cd55f418c1e4a229c961d7a402672309a956965f4b5f52f18bc9270c5d0

Observation 60c842f6-b19c-4219-8770-a2d2879e6b58 · inbound

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks cites this paper.

Robust In-Context Reinforcement Learning Under Reward Poisoning Attacks A Tutorial on Meta-Reinforcement Learning

Reference 2002

Resolution
malformed identifier
no resolver link, observed 2026-08-07T05:52:47.134864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:47.134864Z digest=sha256:c7a3584146f8f1f6dba997a8f89c44ed85332dfc104f8f39e7aba9d041ea9384

Observation 225f2833-5557-4fcc-893b-3bf91fd56b0f · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation A Tutorial on Meta-Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.311817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.311817Z digest=sha256:b923888743ca9b0026ae96c825dd38d74e509500ce98acc0f1ca3090a383e765

Observation 59369bea-951d-4879-922e-ac071fdf0ea3 · inbound

How Should We Meta-Learn Reinforcement Learning Algorithms? cites this paper.

How Should We Meta-Learn Reinforcement Learning Algorithms? A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:48:45.339004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:48:45.339004Z digest=sha256:29d1453981b4806ee06fef5851a1b549c769d89e59483106f5a64f233ccc1983

Observation 683289fe-91dc-4791-8999-9af641b436c9 · inbound

Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control cites this paper.

Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control A Tutorial on Meta-Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:56.168767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:59:56.168767Z digest=sha256:7123c51e9ac6aa7ce7044cfd92f4715fc851068306e107af94918e8a18ae9494

Observation c6e774a5-87ba-4673-8d53-e47ee57090eb · inbound

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism cites this paper.

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:48:51.163438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T01:46:16.923948Z digest=sha256:db34229a809a6b71981fc539be009f0dd2a9555dc642561717f1941474b396ff

Observation 98d02217-19bb-4280-9dea-a4aaf06a2ef1 · inbound

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions cites this paper.

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:31:14.988946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T20:28:46.739742Z digest=sha256:ee52469db93e7a52b349f795f50c82076676db28f24aefe2d84112674d086952

Observation 9abd82dd-6b71-4ffd-90e8-b502cdbc6eed · inbound

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning cites this paper.

Towards Long-Lived Robots: Continual Learning VLA Models via Reinforcement Fine-Tuning A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:10:13.258937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T14:07:10.387869Z digest=sha256:aa0a64bbf51120bd050b71f3f55b4a56e57d6d6e94f4ea8fa8357a39bae9b4ec

Observation 4238129d-7615-4b4d-9c33-b2ff00c67f38 · inbound

Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent cites this paper.

Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent A Tutorial on Meta-Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:46:35.716309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:46:15.275441Z digest=sha256:6a32d636a76b9000463d3af106df226ecc07c558f5f46f93a29bb10be0682578

Observation 363f90f5-c360-4fe4-b15c-2e3d8c7c6886 · inbound

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management cites this paper.

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management A Tutorial on Meta-Reinforcement Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:44:01.443264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T18:42:50.962120Z digest=sha256:7d11d5068cab9a355667611fead3389a5e10daa0a5fe89d3aef2ff3d88c3f5c2

Observation 53776305-ded8-4d2d-94b3-18c786f45477 · inbound

Certificate-Guided Evaluation of Reinforcement Learning Generalization cites this paper.

Certificate-Guided Evaluation of Reinforcement Learning Generalization A Tutorial on Meta-Reinforcement Learning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T20:42:37.754644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T18:23:27.059117Z digest=sha256:5723d37dc8b1d53bee86f0343e5e3ff9dafbaf537f980566646838d440baab9f

Observation 6db16b8c-5030-40f9-b1c8-2c7c846815d2 · inbound

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning cites this paper.

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning A Tutorial on Meta-Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T13:42:27.758405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T13:42:27.758405Z digest=sha256:599b0200cf74979dfeb0b75c8c909fb5c974ebc1341f82126812cea926027800

Observation 1426b008-29c7-463d-8e8c-353fc32af201 · inbound

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models cites this paper.

Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models A Tutorial on Meta-Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.133896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:38:48.912290Z digest=sha256:bac79b1dc2eb5a188981b246d6a8bdc87f7353fc3b2e9cb34ee394f8a11e9acb

Observation e9a06953-df1a-4324-a09c-43cd8d594085 · inbound

Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks cites this paper.

Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks A Tutorial on Meta-Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T19:04:32.205336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T19:04:32.205336Z digest=sha256:93b6fa4f9e37c586bf57416b0181175653dc9ef1e4bdf1f9c31f1fb53c6a4784

Observation 03c40415-f36f-4610-b4ec-c8eb3f89f805 · inbound

ReBRAC-v2: The Return of the King cites this paper.

ReBRAC-v2: The Return of the King A Tutorial on Meta-Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T00:32:39.844330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:32:39.844330Z digest=sha256:b3b381c9a9dbfd337720499d9db47c4674b03a4ae49ec6548c7517609d666e2e