Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:35:59.796489Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2411.11457.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:35:59.796489Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2bdfe559-f9ee-4bae-a6cd-366433b95602 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abb3a62d-6a64-41a0-8675-4821a0df9f8f · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Mishra, S
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a29ac238-2445-4f64-a265-a146a823f43f · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5f06d254-b2f7-44b0-9c53-1105dd80dc89 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8711a58-671b-4cde-9fb0-e2c312369abd · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Learning Relative Return Policies With Upside-Down Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4d3df4d2-cb04-4778-a7ac-f2f997510db1 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Sutton, R
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 356fc9e3-a00e-43f0-8fe1-b43b19252053 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb3827f9-b6af-4954-a12f-8b0c649e1fbb · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7f6b6e94-cfb7-4d46-9236-d7c4e1798352 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d21af92e-d3c8-4ea6-b0ca-c53355df626d · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 805e1cfb-34f3-4478-b0a6-caa6085156b5 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c303bfe3-5043-4dc1-bc5e-94e1c89e15a4 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Hart, P
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0291fb40-5715-4138-b17e-695f96393908 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8b6ea7f-fc98-4de4-9809-652779e16cb0 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5c374c7d-2122-4422-8d62-101f6fee61e1 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e2ae35d7-bfd2-4d5d-b482-327d410a24d3 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation afa1af9c-2172-4f6a-a10f-d1fcd82c4cfb · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control E., et al
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c747cf82-8295-4991-a24c-1f977f7eccfe · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 946b6186-11c2-4f8a-97e6-e932345ed70f · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Generalized Decision Transformer for Offline Hindsight Information Matching
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation becba9f6-d4e7-47b3-b584-eaca283b99b6 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fdef25be-0db2-4f12-a82d-bc3adcb58a4e · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 302823e9-b521-4a8f-919c-2d8d23f4a813 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8ca6e5eb-2068-4c52-aa7a-acf85b2f1f75 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5a6a0b13-8e32-4fc3-b387-3ee9ae08df0d · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 203b2ee3-f44c-41ef-8dac-d0c729cbe1bd · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Pisane, J., Kolios, A., and Ernst, D
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4b876560-7f52-4f3a-b377-c66222a2351a · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Goal-Conditioned Reinforcement Learning: Problems and Solutions
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00a8459a-452e-49f8-934d-8147ed3222b6 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b4402b34-2019-4601-9143-c974431e8000 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4d233def-ac71-4df7-b44a-b982ce9c86c8 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control A., de Lope, J., and Maravall, D
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1a1af645-6662-47f5-99ce-265f2350f2bd · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Q-learning with online random forests
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7ceaf3c9-5804-4fa8-b4d8-1ad46ccde549 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d3a0bc44-5ee8-4a3e-addf-3c405e51ff49 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control J., and Moyà-Alcover, G
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 39603694-50ef-4e27-affb-57c7eb754f85 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0d9da2dd-e1a4-49e1-b420-d93758669cf1 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fcd2e4fa-b7db-4bfe-98e2-fe534b1db697 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4666b078-bf88-4c11-bec8-d7bc8cd7ed62 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d470a829-5ff8-4a2f-8206-074ddbd89a5f · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6724efb-68e3-4366-b7e2-ccf4fcc47c4b · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0c38d81-767c-44ec-ab96-e3a035356c7f · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fce7cf5-155f-4588-9a64-85b414a3262b · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Deep Reinforcement Learning framework for Autonomous Driving
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32f142d3-c417-4019-bb8a-747a4a27778c · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 428271ff-1325-4d6a-a264-5acf5a065854 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Xie, Q
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 78b945d5-c1ce-468f-8318-7bd9334bd59b · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Armon, A
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b5cb02b-cd30-4de8-b819-8532c5117eba · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Wang, L
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0ff21285-d901-4b47-901b-4138ce934c81 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Training Agents using Upside-Down Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c67d7f-c38a-4553-b64b-7f11ccb22ce9 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1ce1ff3d-6701-4dc1-8af7-e86ce1fd4fad · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d4032b09-f706-4d38-8406-cbed089e4fba · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5818cd91-1ac4-4f32-9642-6f5116d0e862 · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 684369a4-35c1-47a2-8e24-5e7eecc5e9eb · outbound
Upside-Down Reinforcement Learning for More Interpretable Optimal Control R., and Zeng, D
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.