Pith. sign in

Paper Citation Record · LEDGER

Upside-Down Reinforcement Learning for More Interpretable Optimal Control

As of 17 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2411.11457.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11457 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:35:59.796489Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved38
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2bdfe559-f9ee-4bae-a6cd-366433b95602 · outbound

This paper cites write newline.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.608437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.608437Z digest=sha256:67834aeae80fa178a694b601f4b57fdb2c2ef8da4b7e44f502209a75c0cb66d5

Observation abb3a62d-6a64-41a0-8675-4821a0df9f8f · outbound

This paper cites and Mishra, S.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Mishra, S

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.399198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.614212Z digest=sha256:f25a6aa63b2af4bc038d582a967f81cf3cc9693743ffbbbec5109a582bc78b78

Observation a29ac238-2445-4f64-a265-a146a823f43f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.387963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.619203Z digest=sha256:04f94f341271ce566c2d308ec1ae9b919a81d209411776f450adf7cb56de0333

Observation 5f06d254-b2f7-44b0-9c53-1105dd80dc89 · outbound

This paper cites All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control All You Need Is Supervised Learning: From Imitation Learning to Meta-RL With Upside Down RL

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.623576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.623576Z digest=sha256:f05c5cb35d1276a81e42a37c868f13fd506515a9790fbe48d3bbf65991a10c36

Observation d8711a58-671b-4cde-9fb0-e2c312369abd · outbound

This paper cites Learning Relative Return Policies With Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Learning Relative Return Policies With Upside-Down Reinforcement Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T18:35:59.922846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.627958Z digest=sha256:6728161e2ecf4c368e76c8770b93671a1713a7f78ca79088c80d990f24d5847e

Observation 4d3df4d2-cb04-4778-a7ac-f2f997510db1 · outbound

This paper cites G., Sutton, R.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Sutton, R

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.375535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.632421Z digest=sha256:5b0dff52d3586aa6dc975b2e7677a4a19fdedc66be6c019da34fae9246b5d5b2

Observation 356fc9e3-a00e-43f0-8fe1-b43b19252053 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.636180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.636180Z digest=sha256:31e3aeec499f0561617aef591e7bc829fa6a7b8be5a57f2df89f4a9fb3fe7e2b

Observation fb3827f9-b6af-4954-a12f-8b0c649e1fbb · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.356335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.640291Z digest=sha256:57522d1a9d052929f9f578dcdc72356316742c642a918676a3407adba1f41736

Observation 7f6b6e94-cfb7-4d46-9236-d7c4e1798352 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.343117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.644021Z digest=sha256:cd7c0bab39b494b13b47749e00539454e50c0568895898af0502db394ed6f890

Observation d21af92e-d3c8-4ea6-b0ca-c53355df626d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.330895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.647621Z digest=sha256:e3828838ed9e8b8260dabba83e15c9f8cea28d01120b2b7f52bbeb01e1dd057d

Observation 805e1cfb-34f3-4478-b0a6-caa6085156b5 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.318334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.651418Z digest=sha256:d8cf7fb9c825ea57d8c490ec4c84ebcbc20332dbabfa09c325e74c9752e86287

Observation c303bfe3-5043-4dc1-bc5e-94e1c89e15a4 · outbound

This paper cites and Hart, P.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Hart, P

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.305489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.656560Z digest=sha256:9a14251f7322311836d2d603253ec281f822b36bcbb33a08e2d4b756cee0a359

Observation 0291fb40-5715-4138-b17e-695f96393908 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.661138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.661138Z digest=sha256:a69af39585b681db431259cd8dc0e44d0728eab408a676afd6a352fa565cf7f8

Observation a8b6ea7f-fc98-4de4-9809-652779e16cb0 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.284611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.664712Z digest=sha256:c0f6b583f135ae476cd705b6329a6d940141362ba751f99a072b563306203ce6

Observation 5c374c7d-2122-4422-8d62-101f6fee61e1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.272324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.668061Z digest=sha256:1ca8409250ce9676118bd94e5703b15efc7f3aef2ef750880698a36ea3c7a5bd

Observation e2ae35d7-bfd2-4d5d-b482-327d410a24d3 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.260146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.671467Z digest=sha256:93b5dcc85a552182f2b0ed72eab4aadc218259a836c922eeffcfbea45688560d

Observation afa1af9c-2172-4f6a-a10f-d1fcd82c4cfb · outbound

This paper cites E., et al.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control E., et al

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.247060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.675111Z digest=sha256:5b7bef0f1cd4290c6e0e05ec40aca54560c61407865e5ad2602fb2a39d72d382

Observation c747cf82-8295-4991-a24c-1f977f7eccfe · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.678651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.678651Z digest=sha256:e5718809b94c1daacab7ed189df20779dfda71df73954fb71d756c00aa4ca696

Observation 946b6186-11c2-4f8a-97e6-e932345ed70f · outbound

This paper cites Generalized Decision Transformer for Offline Hindsight Information Matching.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Generalized Decision Transformer for Offline Hindsight Information Matching

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.682047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.682047Z digest=sha256:13401c2f858dbb4fc31d749e6793061dbfa5000925fd54a29955d82dad0142e2

Observation becba9f6-d4e7-47b3-b584-eaca283b99b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.686181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.686181Z digest=sha256:aec88638e565901383814b128bca79de1c4f4ffbb354dc2f993b187e4d0fd238

Observation fdef25be-0db2-4f12-a82d-bc3adcb58a4e · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.221200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.689567Z digest=sha256:9bf7ca21bdd8adf648bbf04f0a28568c83b3f5f207a036ef73a515e4309a1cbf

Observation 302823e9-b521-4a8f-919c-2d8d23f4a813 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.208953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.692976Z digest=sha256:d826f6bc32abfcaac9346c33ff1dcc3ea558129f58c10e3d6aa42618c5f5a4f8

Observation 8ca6e5eb-2068-4c52-aa7a-acf85b2f1f75 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.197096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.696297Z digest=sha256:d04cd6a57a73080b562788e4f8bba3f632d39b76ba193328cbf72aac227aa4cc

Observation 5a6a0b13-8e32-4fc3-b387-3ee9ae08df0d · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.699667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.699667Z digest=sha256:5a5c2e12f4aa69c4c398a9de8d0158edfb559a27fd80a579ac052fdbf4458f83

Observation 203b2ee3-f44c-41ef-8dac-d0c729cbe1bd · outbound

This paper cites G., Pisane, J., Kolios, A., and Ernst, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control G., Pisane, J., Kolios, A., and Ernst, D

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.176905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.703145Z digest=sha256:0b486ced91bcb5e6280e58c8e658aac5458ab4b86cc3f9a91c79d73ba8eb8d04

Observation 4b876560-7f52-4f3a-b377-c66222a2351a · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.706912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.706912Z digest=sha256:b8acec569ccf0b8ca89148a6b96f9ddf7798a1458b4d3c5727c48eba45cae399

Observation 00a8459a-452e-49f8-934d-8147ed3222b6 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.164422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.711006Z digest=sha256:e577cceaccc54b1302b7bf086d7602d7e23ca03af70bfd451af34aa25a01bbb3

Observation b4402b34-2019-4601-9143-c974431e8000 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.152815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.714854Z digest=sha256:107cf59b66c64ad84994469f66183012e2bcecd1a65ae4e46c080827f90c7ded

Observation 4d233def-ac71-4df7-b44a-b982ce9c86c8 · outbound

This paper cites A., de Lope, J., and Maravall, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A., de Lope, J., and Maravall, D

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.140298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.718392Z digest=sha256:b201ea1dbddbf5407d070f931b727ef67704d6797d24a6b5e759c78cd96d4e6c

Observation 1a1af645-6662-47f5-99ce-265f2350f2bd · outbound

This paper cites Q-learning with online random forests.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Q-learning with online random forests

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:35:59.882991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.722299Z digest=sha256:fd5e06a9044b3693a24a0c54c0e5f269ff0bc922bfaeaf5a2cb48acbd426f799

Observation 7ceaf3c9-5804-4fa8-b4d8-1ad46ccde549 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.128061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.726465Z digest=sha256:ff31ba230163a22af7aa08c7fb24ddaf94888bac730598980813b02563d7e846

Observation d3a0bc44-5ee8-4a3e-addf-3c405e51ff49 · outbound

This paper cites J., and Moyà-Alcover, G.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control J., and Moyà-Alcover, G

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.116517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.730152Z digest=sha256:8ae247afe0817167d964e11caf015dfea03fbaafee220cb3d96c2c5da4a03545

Observation 39603694-50ef-4e27-affb-57c7eb754f85 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.104713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.733683Z digest=sha256:f59dade423242cf307a683e0f4ed507c1ecf9b30d83f0332d899e3a5768c3e81

Observation 0d9da2dd-e1a4-49e1-b420-d93758669cf1 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:36:00.092837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.736993Z digest=sha256:80a039f3218e996a9316b5fe3f54ab293ec1200a634037112e7e4b058a4d9834

Observation fcd2e4fa-b7db-4bfe-98e2-fe534b1db697 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.740227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.740227Z digest=sha256:47126d9d2c649a82dff740e0b03c0a766001b77305e2f188610f45daa43d2609

Observation 4666b078-bf88-4c11-bec8-d7bc8cd7ed62 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.743996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.743996Z digest=sha256:8ddf65014ec69dc007f247ce7e45243af9c049f64ee20cfb849abc2b9306ac48

Observation d470a829-5ff8-4a2f-8206-074ddbd89a5f · outbound

This paper cites A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control A Reinforcement Learning Approach to Weaning of Mechanical Ventilation in Intensive Care Units

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.747396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.747396Z digest=sha256:939a9a5014fdfe86abedc089b6eb073949078db340113561272e074a46a78034

Observation d6724efb-68e3-4366-b7e2-ccf4fcc47c4b · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.751444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.751444Z digest=sha256:fe28bd5bcad594a4270b86df845183f020831c14bafe2c43522a65552d7a24cf

Observation e0c38d81-767c-44ec-ab96-e3a035356c7f · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.755016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.755016Z digest=sha256:7d568153165e7f84b7a74b72853e3c5ba744d7a121ef19129e6d1b3a9064b519

Observation 7fce7cf5-155f-4588-9a64-85b414a3262b · outbound

This paper cites Deep Reinforcement Learning framework for Autonomous Driving.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Deep Reinforcement Learning framework for Autonomous Driving

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.758519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.758519Z digest=sha256:58c557c2072b0612da30bafde4da528cffa5cc986b7fd485154a411bbe3bf6e1

Observation 32f142d3-c417-4019-bb8a-747a4a27778c · outbound

This paper cites Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Reinforcement Learning Upside Down: Don't Predict Rewards -- Just Map Them to Actions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.762452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.762452Z digest=sha256:a04d39c2023ee04fa31a3a754f722c21022b5ac0c46c3b67cc9f530558b71a26

Observation 428271ff-1325-4d6a-a264-5acf5a065854 · outbound

This paper cites and Xie, Q.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Xie, Q

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.048189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.766504Z digest=sha256:27d54fdff5dedfe1be321968d55a58f8546183a7ec305253c6958c1948565f0c

Observation 78b945d5-c1ce-468f-8318-7bd9334bd59b · outbound

This paper cites and Armon, A.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Armon, A

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.771028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.771028Z digest=sha256:aa340dcf65eb741d3f08b376def042add52e91f069d460f911dc8830ad1f7130

Observation 2b5cb02b-cd30-4de8-b819-8532c5117eba · outbound

This paper cites and Wang, L.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control and Wang, L

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:36:00.006631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.774377Z digest=sha256:7d769f455eff4b8469ebee050d61612dcecbbe95c038afe2d0e75288c18d1980

Observation 0ff21285-d901-4b47-901b-4138ce934c81 · outbound

This paper cites Training Agents using Upside-Down Reinforcement Learning.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Training Agents using Upside-Down Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T18:35:59.778062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:35:59.778062Z digest=sha256:78b04ab5995f236005b432cec3fb7f0a5cd09c4ce13e874326f4fb51429511b4

Observation a4c67d7f-c38a-4553-b64b-7f11ccb22ce9 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.992730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.781808Z digest=sha256:14409dd7cc0bbce793287856f7de88a536b986a09febb5d2f61ce27b8d6c4d35

Observation 1ce1ff3d-6701-4dc1-8af7-e86ce1fd4fad · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.981068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.785384Z digest=sha256:4d70b1c2334c25303e0ea50f9cebba8e404b9b1a9d47a285c30f637094654ed1

Observation d4032b09-f706-4d38-8406-cbed089e4fba · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.968977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.789040Z digest=sha256:e953eeb9cf3a6348bb82b6f353164e196cd5238abd593527a3d55552c3090751

Observation 5818cd91-1ac4-4f32-9642-6f5116d0e862 · outbound

This paper cites an unresolved cited work.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T18:35:59.957638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.792631Z digest=sha256:2e8b615a40708023d564cfc67c69f5d8b9714eb72832d11d971ad4e77e65310d

Observation 684369a4-35c1-47a2-8e24-5e7eecc5e9eb · outbound

This paper cites R., and Zeng, D.

Upside-Down Reinforcement Learning for More Interpretable Optimal Control R., and Zeng, D

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:35:59.946338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-12T18:35:59.796489Z digest=sha256:dc95f7e2d36c05fd6176aba03c06c0f1b68631446eb94db0f31aabcdf2fa2889

Pith citing papers

No inbound Pith citation observations are available.