Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2306.14892.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-26T21:50:00.974590Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f4e02161-42d8-43f8-9165-6ef40fa7e2bd · inbound
Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 194
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 51a9fedf-71d2-4fba-a256-aea2cab9baa1 · inbound
One for All: A Non-Linear Transformer can Enable Cross-Domain Generalization for In-Context Reinforcement Learning Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0d394af4-8626-4273-aa97-f8663abe2f41 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c55a68a0-d79c-4972-838b-c8e9b2fef293 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3a18d0c7-fd70-4d9a-9dd7-2e0936c7274d · inbound
Reinforcement Learning Foundation Models Should Already Be A Thing Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e7d950d4-a278-414c-ae1e-03b695a73b5c · inbound
Towards Scalable Multi-Task Reinforcement Learning with Large Decision Models Supervised Pretraining Can Learn In-Context Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.