Pith. sign in

Paper Citation Record · LEDGER

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2506.13672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13672 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:32:54.400821Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact9
  • verified fuzzy21
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 088c0f74-fe5f-4416-898a-0d8fa0a6237d · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.191659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.034330Z digest=sha256:3c175c3a4d0d5cd6c66f03525d35e44c73860bb3c8dbb73a04f93af6307b0a3f

Observation d8740eef-1ef0-4a05-a088-89479bf979f8 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.182111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.070341Z digest=sha256:2f29a294e1b37cb80020db176d6affc178290c0e4d26c64d0b31c6218e1fd692

Observation ba96e2ff-57cd-4c4e-be56-0657917e6209 · outbound

This paper cites OpenAI Gym.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.109869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.109869Z digest=sha256:1bf44f846376006cebfa3772ae7892f075bc610416860cca84a3d95317410fdb

Observation c78b0f73-f7eb-439d-9fa3-82d439d71346 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.172292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.173653Z digest=sha256:4710e906de7458f3786ae3c9df8fe54be370434814775e987c09f9e59579a54e

Observation 94e30e18-13d6-4818-b1e4-163f389d09a6 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.161746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.232671Z digest=sha256:8ad08e853bc9b758567335131616b3ea8d561c831350748e41625b37a0715910

Observation a084e66b-bf25-4ca5-b7ef-e03e92b9180e · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.151864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.284609Z digest=sha256:91f6e83352bd392408bbb1999daf5bd9c9f4fb2d972d24a96e77906a7e40cefc

Observation ac4175c7-aa18-49a8-8509-72ac5a37170c · outbound

This paper cites Stabilizing Off-Policy Deep Reinforcement Learning from Pixels.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Stabilizing Off-Policy Deep Reinforcement Learning from Pixels

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.368740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.368740Z digest=sha256:34f2e21c4b568b24f82373fb64e15e80f8500d14094fd8e005e2040f57e05248

Observation b58b6687-ec6e-4d18-b8aa-6fc75da5b71c · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.412668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.412668Z digest=sha256:e81ba092603bb8e06b35616dd46413db825fc556caf4e819a1ba9fbd9a2e14d6

Observation 47f52821-2c7f-4b10-a141-14f7bd75a73f · outbound

This paper cites Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.773730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.489700Z digest=sha256:0883a72407e62dd4ef41aed2798b3a72312c005d15ce824355da5af2c14001ae

Observation 134041a6-9e97-48bb-8d6f-ba2341c7ff5f · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.925914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.552848Z digest=sha256:cc359d1a5c3bf2d846646f64b5999b97c5a576348ba299673ee00f4900093523

Observation a8db4646-1acd-404e-87ea-702597f337cb · outbound

This paper cites G., and Courville, A.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning G., and Courville, A

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.592830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.592830Z digest=sha256:08d4d0beabd1a31bdabe70ee1717971585cb5e8adf72c026262888953e262e52

Observation 74eb255d-4dae-4ba4-9b9c-1872bf0fcfcd · outbound

This paper cites W., Subramanian, J., and Ghassemi, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., Subramanian, J., and Ghassemi, M

Reference 12

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:32:55.758573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.628887Z digest=sha256:d8717fe5433f774e92e6f39de323cdc539f5efec7824181dcb6f89f4b6ab8a0b

Observation 67b3ad1d-71b8-4f47-973f-7ce034d107c7 · outbound

This paper cites Revisiting fundamentals of experience replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting fundamentals of experience replay

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.703013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.703013Z digest=sha256:d5230470283fdc7b902d3add923e718b5f860711caacf42da2d5d8b502f660fc

Observation b23c158e-46a9-4aa9-b942-a735050d9861 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Addressing function approximation error in actor-critic methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.764866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.764866Z digest=sha256:be3d0ac03eb5c1520f13572f317b622eb6ac4d70d7d4f752199781ceb9bb3f33

Observation c08bb5d9-d454-446f-b43f-21835d58dad6 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.832637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.832637Z digest=sha256:3b725f9a10a09477f576a894153c76deb9014d28eee0e39e3e44a583be41df59

Observation 61054b92-983b-4b0b-831e-0dc697467c82 · outbound

This paper cites Sunk-cost fallacy and cognitive ability in individual decision-making.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sunk-cost fallacy and cognitive ability in individual decision-making

Reference 16

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.706719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.905906Z digest=sha256:9b12fab44aa0beca4a7c423742e32321216fd2da6c89d8a4395663e1fd875ceb

Observation 234a0dae-0fd4-42b5-a57c-451e32a7c932 · outbound

This paper cites R., Millman, K.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning R., Millman, K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.116543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.926428Z digest=sha256:b1244380583dc28117bcf3efd0e6fb3365def9bf02923a0766b84cf9176fe0ce

Observation cf7f6777-c012-435d-ab6e-bb8e2237927d · outbound

This paper cites Double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Double q-learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.107696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.025987Z digest=sha256:55f67e1bc86d255bb5400800a64a9c87fb103cb7e1c7cf339380d05917fbf798

Observation 962346cb-77aa-4e31-a6e4-c439b653c794 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.089069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.089069Z digest=sha256:9edce17f669649ca5fe24ff49195d8e0394bfa8771c6ba66cb6eb095af04ba7e

Observation 30c170d8-8f14-4a6c-8741-82a3045d317f · outbound

This paper cites Planning Goals for Exploration.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Planning Goals for Exploration

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.133496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.133496Z digest=sha256:4b3f6ac107f781d26cce4fce9d9b26b5bf4ae4b0f533c43837d3881ab69fe489

Observation 0b6767b6-f2ed-4dfd-8933-e080872a10b3 · outbound

This paper cites Enhanced Experience Replay Generation for Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Enhanced Experience Replay Generation for Efficient Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.201826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.201826Z digest=sha256:5e27d9f042629dc025b3e0f95621ee3daf93fd61372f4dd4a1bd758bc362bbd3

Observation a7febbd2-a0e6-47f6-8055-a046938b0574 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.098144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.236555Z digest=sha256:a4172ffd405449e4b095d7fee49ef1d8b2cbe68c8f1ded539a5786b8d42bf6ec

Observation a28efe13-186d-491b-926d-7db7514d8bb1 · outbound

This paper cites An investigation of generative replay in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning An investigation of generative replay in deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.089516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.264341Z digest=sha256:24a28da94b8ee113460df76457bec99702406e235e81d7b588a7fdac7dcae97d

Observation f8429033-e01f-4244-9be0-55bba015ae5e · outbound

This paper cites Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:32:55.635890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.310272Z digest=sha256:94124ffa60246fd420caf39e9ae936ffe81b24e5cf78ec100f7723ab81a0cbf0

Observation f931b24d-da63-42ea-8c8b-e07459ba2322 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.372728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.372728Z digest=sha256:812456ea900a62cc2d6bde9931ce4f97217c84459e433bc4764bad688937af4d

Observation 7f149a22-56a5-47e0-b243-32fd713826ff · outbound

This paper cites CURL : Contrastive unsupervised representations for reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.080060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.422423Z digest=sha256:3080b3254cce3e45fa9f325c72602a4707b37b4edceeac6af6b55103bac8efa3

Observation 4eba7925-a8d6-498a-bb0c-2245d0a24203 · outbound

This paper cites HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.484896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.484896Z digest=sha256:6c64a2174130e304d22ec1f0a06e4c9704406f050d79cac50ae54c1c8e2c8afc

Observation 9027ad0e-f8f2-4657-a7be-1ecfbd57f57a · outbound

This paper cites Continuous control with deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Continuous control with deep reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.581844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.581844Z digest=sha256:26cf398262e08f4352d6a52686010b4ca87b9ba60a6f77a8ac7918539de8d42b

Observation 8ca1f07c-c3b9-4933-8c83-ec8274f9d30f · outbound

This paper cites Unlock the intermittent control ability of model free reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unlock the intermittent control ability of model free reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.069185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.627229Z digest=sha256:395e3ef6497577daa71f34d9b0cb7819b9ad1f31b0a2c26939b9beadccc5bbd6

Observation 66a22057-2102-44d2-91ba-467c9ce178c7 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.060067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.676803Z digest=sha256:5fdabb51e4649bb8a7bf6d1c518a2885e6025c910afe01585b3b1836be58b024

Observation 777ddf33-8dd3-45f6-841b-b27fa644f0c7 · outbound

This paper cites W., and Parker-Holder, J.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., and Parker-Holder, J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.049548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.718788Z digest=sha256:2ca290e94868f18d734476bb8db697e580bc1e9f9ecf54c15b5c86de8b230b98

Observation 5a3c5aca-0cec-4191-a7ba-ccd0c65ce5ea · outbound

This paper cites Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.760568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.760568Z digest=sha256:ad7ccf04818d483b8d4e815ace88dc23f9e769b75c54b84b47195ac47669be79

Observation 914aae6e-5090-4352-802d-356c8f4d15ef · outbound

This paper cites Online reinforcement learning with uncertain episode lengths.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Online reinforcement learning with uncertain episode lengths

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.040234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.823709Z digest=sha256:1df6e76121cd9f54d4334d16837f6494aaeed9ddb9e2b0fd996c2215728a6fda

Observation 1bada11b-e190-455f-b112-7defdb9beb3d · outbound

This paper cites Tactical optimism and pessimism for deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Tactical optimism and pessimism for deep reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.029700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.877004Z digest=sha256:4a706e5e4793c60a7fb4a7beb2b0947198355b11d5e918b3614f1a5e58d20ac1

Observation 0d1d9e27-5aed-4b82-a953-88d50f1a3aeb · outbound

This paper cites Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.919314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.919314Z digest=sha256:e2b55d8d1132d0349306ada6dfb778558c4bd9e4f7fcb559bf7a413da8247314

Observation 070c5cfe-e979-4beb-bcc9-a55dbd64ec20 · outbound

This paper cites The primacy bias in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning The primacy bias in deep reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.016950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.968328Z digest=sha256:4f1d34610a94d0a08791f2fdc0b6741b407cf5d4acab9304879ce8e623ea2ca2

Observation e8eb03d9-2105-45a2-bbaa-990e2397a778 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.017309Z digest=sha256:bbc9a84f34eaa663d61231c33a9c645e3205857ad363e442efa2723da2ae6be4

Observation c0fc6932-682d-482a-bfc2-4920c266cc36 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.992729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.051990Z digest=sha256:9ae34702b785b91b0c3ec989ceba8ae4a84339243bbd57429362865e5c689b83

Observation 8668d5d8-ef85-4d7b-8892-2c64c2f76568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.099844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.099844Z digest=sha256:2b03161c95e15bcc771eaf93af55b2c8db7473940c092da859c3f091bec2a444

Observation 65a0b996-78b1-4ed1-a3ea-e92db6cdf013 · outbound

This paper cites Reinforcement Learning with Dynamic Boltzmann Softmax Updates.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Reinforcement Learning with Dynamic Boltzmann Softmax Updates

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.160250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.160250Z digest=sha256:00981fe95a68dbb39edf091f4c716b607a39d3141e0cc4e8efb906c770f575d9

Observation 9660b867-6b68-419b-88d0-fd77ae2cf9f2 · outbound

This paper cites Softmax deep double deterministic policy gradients.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Softmax deep double deterministic policy gradients

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.981193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.192637Z digest=sha256:25857e70d2f70f02f66591cd2b8823e9042818a7906950261a166ba9c5b87b07

Observation fd8f489c-b566-440a-88c7-cf0253156be3 · outbound

This paper cites Regularized softmax deep multi-agent q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Regularized softmax deep multi-agent q-learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.970541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.230814Z digest=sha256:7fdb33749003508d834e6eb41ab15cae8c9715ebd561eef783c42bd4712b420b

Observation 5c1641d6-a350-4942-84a0-4fc0f58aa313 · outbound

This paper cites Time limits in reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Time limits in reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.958785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.321621Z digest=sha256:6516f1e45273c719bd3f4ea692139d8aa9b3426e68963e7eda2a9fe4180f52e4

Observation a39b91ca-6935-48b4-a1b8-47d17ec83dd5 · outbound

This paper cites A., and Darrell, T.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning A., and Darrell, T

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.947216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.386837Z digest=sha256:bda5c9a41b18cbb3cf1d273b2f7f2d2550e06cde71f95a6255dde853b599b67d

Observation 2986c97b-8c8d-4f18-b235-2c43b089d4f4 · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.934747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.430420Z digest=sha256:2422066db90cf7b5c2a3a77aef6dcd76941221b96320956e80942eaaef021517

Observation 42e3395c-bc25-4354-8cb0-42c32b25495c · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.921337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.493835Z digest=sha256:a905acb1f92f246d178ca31da74bebe3a6ce49e53d0c601264236f53801a9674

Observation ac7e3916-194d-49ab-9420-f5302e4f0da4 · outbound

This paper cites Optimistic Exploration even with a Pessimistic Initialisation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimistic Exploration even with a Pessimistic Initialisation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.555362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.586211Z digest=sha256:fb571ea1aa9df7deaee1fe26820cad4697104637d836cf2b8daddd80096a9d4a

Observation b7cee4a3-ece9-4414-85ac-3e7ec8f782a7 · outbound

This paper cites Prioritized Experience Replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.629886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.629886Z digest=sha256:78231036b60fecfc6c2ad1abd4d11ca05b9ac547b502364e0ad9d310b5d54c88

Observation 0e458a88-e4c5-4099-9924-1a5abfd97568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.908081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.697565Z digest=sha256:f449396e7446c426ec31bd9e5c273972abe02b522f6ab166703c3edc55236fce

Observation e9e3441b-e682-48b6-9f1b-4ea00bcb32b4 · outbound

This paper cites S., and Evci, U.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning S., and Evci, U

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.837429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.837429Z digest=sha256:4a942daa99da6f2cb4ac299ee87e923834ac8ba230f5e8cc49299dd5af77ef2b

Observation 6a35abc8-8d00-4b36-893e-a38fb2906a11 · outbound

This paper cites Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.490510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.902691Z digest=sha256:62e4ab8c82e3e446175759a9d5e76dfd61e52fc4affe4dd0818cb36b49cd1147

Observation f4ab1462-9de7-4f2a-8679-bc0723435502 · outbound

This paper cites Revisiting the softmax bellman operator: New benefits and new perspective.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting the softmax bellman operator: New benefits and new perspective

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.886519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.972733Z digest=sha256:e91f50febf9aa63c16f2b11cecad5ae0fdaa58ace992eebc20054c480f4f5266

Observation d90db58a-4793-4d0b-9c9c-9099759ce176 · outbound

This paper cites Prioritizing samples in reinforcement learning with reducible loss.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritizing samples in reinforcement learning with reducible loss

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.876288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.084002Z digest=sha256:cbe588e56f3c98c9cf25840e576697db960123ba65b7dafba04e94e764dd42ac

Observation e6907165-b461-4f4a-ad31-8153b8e506b7 · outbound

This paper cites Safe Exploration by Solving Early Terminated MDP.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Safe Exploration by Solving Early Terminated MDP

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.325475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.197634Z digest=sha256:d23404fc75a7fd70c165be71e69fbe9c0275b77ccded6b0b1a23e56c5be04aae

Observation 35d61335-ae2b-479d-abab-982bc9a399e0 · outbound

This paper cites sunk costs.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning sunk costs

Reference 56

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.587105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.278813Z digest=sha256:0d7546cbd657c5878599b7357b7434fad0662977db76c8f54f804d62364ae809

Observation 364092c5-6388-4b31-afdc-14272e675950 · outbound

This paper cites DeepMind Control Suite.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DeepMind Control Suite

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.351761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.351761Z digest=sha256:5d606b648f8cc1f6e7a7e8b89930afec0aeb4acdef1acead806dfbf51191183d

Observation 6e83b4c8-5af9-4201-9939-05fa456612a6 · outbound

This paper cites Loss Functions and Metrics in Deep Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Loss Functions and Metrics in Deep Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.432291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.432291Z digest=sha256:94daaec3bf2acd5db43a5f7453b53e3fee1f2f054b7c1d6bc1a328e89aeda0e9

Observation 388bbcb7-8741-43ce-a4a1-fe19799cd702 · outbound

This paper cites H., Meyers, E.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning H., Meyers, E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.866665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.623704Z digest=sha256:3ef8c01fc1cc722215d27dd9088ef15d1660eda8ccfbfd0ba5e93bc091db9dc1

Observation 31640fed-96f4-4f9b-9cd6-58fb38c952ce · outbound

This paper cites Deep reinforcement learning with double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Deep reinforcement learning with double q-learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.857507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.697176Z digest=sha256:531de39cf37817a19a09caa360be4671db0ce6ee33e1fc6e27096f6b8056dea7

Observation 58976586-1628-4cf5-a242-d26288add4d0 · outbound

This paper cites and Drake Jr, F.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning and Drake Jr, F

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.847097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.713465Z digest=sha256:f3488c17aaf7d9cb1cd87cee3ac83d0c98b2971019d71740307112ae4c1bab66

Observation cc25d49c-21eb-4831-96f8-e5b9a1915127 · outbound

This paper cites Optimism in Reinforcement Learning with Generalized Linear Function Approximation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimism in Reinforcement Learning with Generalized Linear Function Approximation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.761065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.761065Z digest=sha256:6296bea123261b991c5c6110b9a75c74bc2794a43a87ea1ae60c23e5e6e8f3f0

Observation 0f10747d-6b07-4af0-b75b-552092ec3032 · outbound

This paper cites DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.901414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.901414Z digest=sha256:8f44cfa54853cc45566ef3b81660394f157ee76f1a30f8cab458f5774709d546

Observation c727a42e-9e6c-4af3-97cf-7fda9d49554c · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.985431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.985431Z digest=sha256:9777f783b7014b5745dec0b01d717345f2531eedfd9c366ccc4cc74bf8a6b11f

Observation a72e8898-159d-41b9-a619-b35c1687e2dc · outbound

This paper cites Sample Efficient Deep Reinforcement Learning via Local Planning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sample Efficient Deep Reinforcement Learning via Local Planning

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.136783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.102945Z digest=sha256:67e0ac869c108115064dd086a263632751d8a20d2371aac81edd2437a9e91647

Observation 0475043e-986d-4206-bde3-a48ea046d7fb · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Scaling Robot Learning with Semantically Imagined Experience

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.190932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.190932Z digest=sha256:5c7eb48f4f9d26494860d4a713c9eda29cd9f4a96d61d38de7f262be0ba47b0c

Observation 0b575bae-4391-42df-9f0a-3a9a92463523 · outbound

This paper cites Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.835270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.287162Z digest=sha256:3bc438a8d67ebe3344ece04c444eb70721d94bd81a5b9fcaa7ab683346de0cef

Observation ea620ca8-f979-4a01-ac43-e0762f4be893 · outbound

This paper cites write newline.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning write newline

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.400821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.400821Z digest=sha256:4ea7e6dcbbd97c83eed81ab75ac944ba933c15bb1b5bfc45dfcbf0eddc3571bd

Pith citing papers

No inbound Pith citation observations are available.