Pith. sign in

Paper Citation Record · LEDGER

Upside-Down Reinforcement Learning for More Interpretable Optimal Control

As of 20 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2411.11457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11457 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:35:59.796489Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2bdfe559-f9ee-4bae-a6cd-366433b95602 · outbound

This paper cites write newline.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.608437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.608437Z digest=sha256:67834aeae80fa178a694b601f4b57fdb2c2ef8da4b7e44f502209a75c0cb66d5

Observation abb3a62d-6a64-41a0-8675-4821a0df9f8f · outbound

This paper cites and Mishra, S.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Mishra, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.399198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.614212Z digest=sha256:9f5474db36fc70c9a2455c4312c5144a6953eb91ba7ef26e532eca189d8df73b

Observation a29ac238-2445-4f64-a265-a146a823f43f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.387963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.619203Z digest=sha256:bccb9f42dcd64ecff97e469c83a1eb57f732b01090e32b0d594abbcb9685e257

Observation 5f06d254-b2f7-44b0-9c53-1105dd80dc89 · outbound

This paper cites All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.623576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.623576Z digest=sha256:f05c5cb35d1276a81e42a37c868f13fd506515a9790fbe48d3bbf65991a10c36

Observation d8711a58-671b-4cde-9fb0-e2c312369abd · outbound

This paper cites Learning Relative Return Policies With Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Learning Relative Return Policies With Upside-Down Reinforcement Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T18:35:59.922846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.627958Z digest=sha256:5335715a874b740da0736b153387f0bf37da1a2cedc95fc47f21f06f736c40c5

Observation 4d3df4d2-cb04-4778-a7ac-f2f997510db1 · outbound

This paper cites G., Sutton, R.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Sutton, R

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.375535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.632421Z digest=sha256:4d97f07cd867aeea4b87c34e47e0c8a201f1140cab58454a79f2ef55c1cb6373

Observation 356fc9e3-a00e-43f0-8fe1-b43b19252053 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.636180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.636180Z digest=sha256:31e3aeec499f0561617aef591e7bc829fa6a7b8be5a57f2df89f4a9fb3fe7e2b

Observation fb3827f9-b6af-4954-a12f-8b0c649e1fbb · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.356335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.640291Z digest=sha256:3213533676b5b2df528dfea5e667a113a6de189e040b42121e599fbcb2e4cae6

Observation 7f6b6e94-cfb7-4d46-9236-d7c4e1798352 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.343117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.644021Z digest=sha256:40fcd3382d74dc1680cc60ddfc9d1edb9a89ff76739bd0798c65ded8eae04f6c

Observation d21af92e-d3c8-4ea6-b0ca-c53355df626d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.330895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.647621Z digest=sha256:0ce058fc3d288b361c74ab6885d1f79ecbbe9c2bee6750888a570b4343198410

Observation 805e1cfb-34f3-4478-b0a6-caa6085156b5 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.318334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.651418Z digest=sha256:0232c0d6b5e949a6a4dd07175ae7be2aabd1dc1c539f1b8402d2f12f76d6a1c1

Observation c303bfe3-5043-4dc1-bc5e-94e1c89e15a4 · outbound

This paper cites and Hart, P.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Hart, P

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.305489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.656560Z digest=sha256:3a1059acb475d62bdd7d9256cdb35b2011c276aa3167adb2af436098dfea2002

Observation 0291fb40-5715-4138-b17e-695f96393908 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.661138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.661138Z digest=sha256:a69af39585b681db431259cd8dc0e44d0728eab408a676afd6a352fa565cf7f8

Observation a8b6ea7f-fc98-4de4-9809-652779e16cb0 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.284611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.664712Z digest=sha256:ffed456fea5ed9caaf828aca7bf5ab5c3276753785902f5d8b1e955789fca9bb

Observation 5c374c7d-2122-4422-8d62-101f6fee61e1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.272324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.668061Z digest=sha256:af8348ec5dd10d3ae60ab724e5bd9c888d802d0127d498dd98cc6185906605de

Observation e2ae35d7-bfd2-4d5d-b482-327d410a24d3 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.260146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.671467Z digest=sha256:874df7c8431524691a659e483f3b259f604c21b2a68d8d2c11048f85801afcef

Observation afa1af9c-2172-4f6a-a10f-d1fcd82c4cfb · outbound

This paper cites E., et al.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control E., et al

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.247060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.675111Z digest=sha256:3ad58af195ef0be7877cbc160395dfadec7336baf4ffc0f11b521bd3897d56ab

Observation c747cf82-8295-4991-a24c-1f977f7eccfe · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.678651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.678651Z digest=sha256:e5718809b94c1daacab7ed189df20779dfda71df73954fb71d756c00aa4ca696

Observation 946b6186-11c2-4f8a-97e6-e932345ed70f · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.682047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.682047Z digest=sha256:13401c2f858dbb4fc31d749e6793061dbfa5000925fd54a29955d82dad0142e2

Observation becba9f6-d4e7-47b3-b584-eaca283b99b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.686181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.686181Z digest=sha256:aec88638e565901383814b128bca79de1c4f4ffbb354dc2f993b187e4d0fd238

Observation fdef25be-0db2-4f12-a82d-bc3adcb58a4e · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.221200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.689567Z digest=sha256:0a49aea6372e64c04f1ca63faf535486d15b18dee6856f933106bde4f923d513

Observation 302823e9-b521-4a8f-919c-2d8d23f4a813 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.208953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.692976Z digest=sha256:3fb47610b94400d4f6861f249cd1b672adb56668d03e0461f0e180dc5035c997

Observation 8ca6e5eb-2068-4c52-aa7a-acf85b2f1f75 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.197096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.696297Z digest=sha256:20e612041b2a0daa578ef590c0e6355c53b5020ebf9c237a0d235bfbc0dd9e9a

Observation 5a6a0b13-8e32-4fc3-b387-3ee9ae08df0d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.699667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.699667Z digest=sha256:5a5c2e12f4aa69c4c398a9de8d0158edfb559a27fd80a579ac052fdbf4458f83

Observation 203b2ee3-f44c-41ef-8dac-d0c729cbe1bd · outbound

This paper cites G., Pisane, J., Kolios, A., and Ernst, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Pisane, J., Kolios, A., and Ernst, D

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.176905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.703145Z digest=sha256:b7285881a6e7fd2d0e00c683ecfa9b59ede5a20727baf62cd03da397a3a3296b

Observation 4b876560-7f52-4f3a-b377-c66222a2351a · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.706912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.706912Z digest=sha256:b8acec569ccf0b8ca89148a6b96f9ddf7798a1458b4d3c5727c48eba45cae399

Observation 00a8459a-452e-49f8-934d-8147ed3222b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.164422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.711006Z digest=sha256:9404f8c38eaa502d8e2d2acf14d381a299d8b01cb4a8609432c3de28ee85a34e

Observation b4402b34-2019-4601-9143-c974431e8000 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.152815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.714854Z digest=sha256:0165addabbcf1943f74a3172bfa64c70feb438ea18275c74e7651134f10bcfc5

Observation 4d233def-ac71-4df7-b44a-b982ce9c86c8 · outbound

This paper cites A., de Lope, J., and Maravall, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A., de Lope, J., and Maravall, D

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.140298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.718392Z digest=sha256:bb32043fc0ea8109bb472c9da4af6b9358eda7bf4b70e06ffc864d0d04da1c5a

Observation 1a1af645-6662-47f5-99ce-265f2350f2bd · outbound

This paper cites Q-learning with online random forests.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Q-learning with online random forests

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:35:59.882991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.722299Z digest=sha256:1272d870ea22aa112feed145d7db52427cb5c42e59117f258d1af8c9094e9dcf

Observation 7ceaf3c9-5804-4fa8-b4d8-1ad46ccde549 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.128061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.726465Z digest=sha256:f941a0185193e41fdaee05f9a5429f3d340572453fcc81fef5ac73d5d7c94816

Observation d3a0bc44-5ee8-4a3e-addf-3c405e51ff49 · outbound

This paper cites J., and Moyà-Alcover, G.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control J., and Moyà-Alcover, G

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.116517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.730152Z digest=sha256:b5703126a45c6586ee56339f0cdd717364ea3cb5910687e588f723ca212b4fb4

Observation 39603694-50ef-4e27-affb-57c7eb754f85 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.104713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.733683Z digest=sha256:3fab8084d32a6cd5536a86bf8c74a8291f95b7e2b8395d262a8ca44c519a4f38

Observation 0d9da2dd-e1a4-49e1-b420-d93758669cf1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.092837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.736993Z digest=sha256:d48cd3fba4011e4d21a32969a4eb2e1de2bb022cae5679b4c2c26a263d307dfa

Observation fcd2e4fa-b7db-4bfe-98e2-fe534b1db697 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.740227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.740227Z digest=sha256:47126d9d2c649a82dff740e0b03c0a766001b77305e2f188610f45daa43d2609

Observation 4666b078-bf88-4c11-bec8-d7bc8cd7ed62 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.743996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.743996Z digest=sha256:8ddf65014ec69dc007f247ce7e45243af9c049f64ee20cfb849abc2b9306ac48

Observation d470a829-5ff8-4a2f-8206-074ddbd89a5f · outbound

This paper cites A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.747396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.747396Z digest=sha256:939a9a5014fdfe86abedc089b6eb073949078db340113561272e074a46a78034

Observation d6724efb-68e3-4366-b7e2-ccf4fcc47c4b · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.751444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.751444Z digest=sha256:fe28bd5bcad594a4270b86df845183f020831c14bafe2c43522a65552d7a24cf

Observation e0c38d81-767c-44ec-ab96-e3a035356c7f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.755016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.755016Z digest=sha256:7d568153165e7f84b7a74b72853e3c5ba744d7a121ef19129e6d1b3a9064b519

Observation 7fce7cf5-155f-4588-9a64-85b414a3262b · outbound

This paper cites Deep Reinforcement Learning framework for Autonomous Driving.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Deep Reinforcement Learning framework for Autonomous Driving

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.758519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.758519Z digest=sha256:82749f2409d54b6869407f5634db482060ebbda77de27bddf107a18f7dd19b47

Observation 32f142d3-c417-4019-bb8a-747a4a27778c · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.762452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.762452Z digest=sha256:a04d39c2023ee04fa31a3a754f722c21022b5ac0c46c3b67cc9f530558b71a26

Observation 428271ff-1325-4d6a-a264-5acf5a065854 · outbound

This paper cites and Xie, Q.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Xie, Q

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.048189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.766504Z digest=sha256:71bfa30f2f0307df117fa453e3a2984fec0a2925869c01bdd97e63a4bf686a55

Observation 78b945d5-c1ce-468f-8318-7bd9334bd59b · outbound

This paper cites and Armon, A.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Armon, A

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.771028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.771028Z digest=sha256:aa340dcf65eb741d3f08b376def042add52e91f069d460f911dc8830ad1f7130

Observation 2b5cb02b-cd30-4de8-b819-8532c5117eba · outbound

This paper cites and Wang, L.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Wang, L

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.006631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.774377Z digest=sha256:89400d7e1d329071e8e64c337686a96409ed912009f726e42874fa50167402f4

Observation 0ff21285-d901-4b47-901b-4138ce934c81 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Training Agents using Upside-Down Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.778062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.778062Z digest=sha256:78b04ab5995f236005b432cec3fb7f0a5cd09c4ce13e874326f4fb51429511b4

Observation a4c67d7f-c38a-4553-b64b-7f11ccb22ce9 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.992730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.781808Z digest=sha256:69718fd7e68e8ec9e3c9095c05a80d693948bf216eee6d8ea0e8db07b9792dc0

Observation 1ce1ff3d-6701-4dc1-8af7-e86ce1fd4fad · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.981068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.785384Z digest=sha256:67cb7c32804c611714974d378848176e8dc7dca70615a879fde442066658e136

Observation d4032b09-f706-4d38-8406-cbed089e4fba · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.968977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.789040Z digest=sha256:257f7e3c10e9f95a515cfe7641cf74c300754edac8375c6f868100168330c7ec

Observation 5818cd91-1ac4-4f32-9642-6f5116d0e862 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.957638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.792631Z digest=sha256:1620bc4211066fa9498f4a37bb740cc946b7662a4a795be05767b93a61cc155c

Observation 684369a4-35c1-47a2-8e24-5e7eecc5e9eb · outbound

This paper cites R., and Zeng, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control R., and Zeng, D

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:35:59.946338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.796489Z digest=sha256:471faf1c2688ec6263336f4897dee0290ebd5acf65447279a2aa090b3d32063a

Pith citing papers

No inbound Pith citation observations are available.