Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:1912.02877.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:30:24.270417Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T21:57:25.915173Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 80bf680b-5b57-4d18-8841-ae1b5076cacc · inbound
Decision Transformer: Reinforcement Learning via Sequence Modeling Training Agents using Upside-Down Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1baf92be-02ec-4bbb-be46-cfa2500d5d2e · inbound
Is Conditional Generative Modeling all you need for Decision-Making? Training Agents using Upside-Down Reinforcement Learning
Reference 209
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38f45699-690a-442d-9483-3fb582ba6590 · inbound
MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework Training Agents using Upside-Down Reinforcement Learning
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc5c38cc-235b-47fc-a44a-eabe153027c2 · inbound
A Provable Approach for End-to-End Safe Reinforcement Learning Training Agents using Upside-Down Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e69c1d7-7926-44ab-b3ac-bb44ef679d6c · inbound
How to Provably Improve Return Conditioned Supervised Learning? Training Agents using Upside-Down Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 325a3a7e-8201-4ff1-971f-797eabf7a0e8 · inbound
Behavioral Exploration: Learning to Explore via In-Context Adaptation Training Agents using Upside-Down Reinforcement Learning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 715459ac-bfe6-476e-b793-aea24c7a4984 · inbound
Equivariant Goal Conditioned Contrastive Reinforcement Learning Training Agents using Upside-Down Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28fa6eef-fb21-4a98-897d-7278f3b6ea36 · inbound
GeoExplorer: Active Geo-localization with Curiosity-Driven Exploration Training Agents using Upside-Down Reinforcement Learning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 310d9c01-bbdb-4460-8592-6e9a51c52323 · inbound
Success Conditioning as Policy Improvement: The Optimization Problem Solved by Imitating Success Training Agents using Upside-Down Reinforcement Learning
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d1701b-aded-4818-9ad6-1aa0f6fc72eb · inbound
QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Training Agents using Upside-Down Reinforcement Learning
Reference 232
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b33bb334-208a-4f7b-b6bd-9a6d9280bba5 · inbound
Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies Training Agents using Upside-Down Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fc21b51e-44c7-4011-86bc-1d869c0540fc · inbound
Reinforcement Learning: From Algorithms To Foundation Models Training Agents using Upside-Down Reinforcement Learning
Reference 174
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 008ffa89-a4da-4946-90dd-a2c1f29c159c · inbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Training Agents using Upside-Down Reinforcement Learning
Reference 269
Source-reported events for the cited work
Unavailable: canonical work link unavailable.