Pith. sign in

Paper Citation Record · LEDGER

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2506.13672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13672 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:32:54.400821Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact9
  • verified fuzzy21
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 088c0f74-fe5f-4416-898a-0d8fa0a6237d · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.191659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.034330Z digest=sha256:b713d70cbbe968d6118ca183bbedff9da48b04b1f6ccecbae50624e47d8ad7a5

Observation d8740eef-1ef0-4a05-a088-89479bf979f8 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.182111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.070341Z digest=sha256:651fae1240d6706af992a965285519071e2418205025b336a928ab421aed6392

Observation ba96e2ff-57cd-4c4e-be56-0657917e6209 · outbound

This paper cites OpenAI Gym.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.109869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.109869Z digest=sha256:872f51d2a66e1df59e10acfae7a1493a40628df2f3c79a058f2588528fc81a2e

Observation c78b0f73-f7eb-439d-9fa3-82d439d71346 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.172292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.173653Z digest=sha256:23dff41bc972efe5d12920c1b5f241cfcabc06ad459cf9846b2c4c2e5e4a47a6

Observation 94e30e18-13d6-4818-b1e4-163f389d09a6 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.161746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.232671Z digest=sha256:3fca2bb35275338f9036b391022cdcc13a07ea961196baa8763dadfe634b5b99

Observation a084e66b-bf25-4ca5-b7ef-e03e92b9180e · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.151864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.284609Z digest=sha256:0ec5dd950287ad0bada379ed46564a49acbb288935f720dcec8fc0114b6f921f

Observation ac4175c7-aa18-49a8-8509-72ac5a37170c · outbound

This paper cites Stabilizing Off-Policy Deep Reinforcement Learning from Pixels.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Stabilizing Off-Policy Deep Reinforcement Learning from Pixels

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.368740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.368740Z digest=sha256:3e8f6343b942c61e455741603233ddce54280d705a16a870d8d5125d443c43d7

Observation b58b6687-ec6e-4d18-b8aa-6fc75da5b71c · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.412668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.412668Z digest=sha256:9e7250a0d1af07a0d12a317bd6106ddcd353cdb9fbeacef6b4471636ccb5d3c0

Observation 47f52821-2c7f-4b10-a141-14f7bd75a73f · outbound

This paper cites Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.773730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.489700Z digest=sha256:9cb14bed212dcc1b4f2fcf11f168311c3682910aa80fc57e55c502f6dcfa0e3c

Observation 134041a6-9e97-48bb-8d6f-ba2341c7ff5f · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.925914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.552848Z digest=sha256:65e41471cd34d913c2c8f95d1107efe384c1a7adbbc92cd120aa7b87238bc90d

Observation a8db4646-1acd-404e-87ea-702597f337cb · outbound

This paper cites G., and Courville, A.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning G., and Courville, A

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.592830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.592830Z digest=sha256:e57124b1f85f32fb611e3b5a8cdfa8155f99b86026500abfdf6ef1edfc050508

Observation 74eb255d-4dae-4ba4-9b9c-1872bf0fcfcd · outbound

This paper cites W., Subramanian, J., and Ghassemi, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., Subramanian, J., and Ghassemi, M

Reference 12

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:32:55.758573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.628887Z digest=sha256:be4658bafcc80aa83ec45d57fd50295e1693e579b1ad2de61548807265148325

Observation 67b3ad1d-71b8-4f47-973f-7ce034d107c7 · outbound

This paper cites Revisiting fundamentals of experience replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting fundamentals of experience replay

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.703013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.703013Z digest=sha256:3dc8020efa20a99a6cd8e5b59cd072d9cf79f4fe563ae24b7fc61d234abe3def

Observation b23c158e-46a9-4aa9-b942-a735050d9861 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Addressing function approximation error in actor-critic methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.764866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.764866Z digest=sha256:a5d006b87083265f5d2340176b7f0972a797e82a4ce531b9e7856978809dc27a

Observation c08bb5d9-d454-446f-b43f-21835d58dad6 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.832637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.832637Z digest=sha256:2512fd26703cf072f0f8afb394b45809332920bd4f9001b1a37c3f2de7153e80

Observation 61054b92-983b-4b0b-831e-0dc697467c82 · outbound

This paper cites Sunk-cost fallacy and cognitive ability in individual decision-making.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sunk-cost fallacy and cognitive ability in individual decision-making

Reference 16

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.706719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.905906Z digest=sha256:1e2300b38f784c5c2086722dd2c19ab1cdc32fa0a54fec68ca4b2c58e4af6eb2

Observation 234a0dae-0fd4-42b5-a57c-451e32a7c932 · outbound

This paper cites R., Millman, K.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning R., Millman, K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.116543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.926428Z digest=sha256:19c36ded326b30f7ac5e9e1c59d5d570962906f0e5d266c67bf46f7f3fee4532

Observation cf7f6777-c012-435d-ab6e-bb8e2237927d · outbound

This paper cites Double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Double q-learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.107696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.025987Z digest=sha256:2dbf2d252f261fc9251be21beb02143d763412bad5c2f212c66536a291244319

Observation 962346cb-77aa-4e31-a6e4-c439b653c794 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.089069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.089069Z digest=sha256:ecb16aed1eb49833d034909b970eb3a640b8e57cc0f2289942d2f48f99ec41e8

Observation 30c170d8-8f14-4a6c-8741-82a3045d317f · outbound

This paper cites Planning Goals for Exploration.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Planning Goals for Exploration

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.133496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.133496Z digest=sha256:363c702b7abadceb3ab7948dcc771c7b6e0480b83229c369ed6f1eccfa6a91c5

Observation 0b6767b6-f2ed-4dfd-8933-e080872a10b3 · outbound

This paper cites Enhanced Experience Replay Generation for Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Enhanced Experience Replay Generation for Efficient Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.201826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.201826Z digest=sha256:507316a3bd1b47f04c07ede61624f27a407667aef42115ade777c1a4f251d31f

Observation a7febbd2-a0e6-47f6-8055-a046938b0574 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.098144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.236555Z digest=sha256:97f593bb9fffbb52a980ea86ec0518fda1b5c762c3454541c24f3238b760d473

Observation a28efe13-186d-491b-926d-7db7514d8bb1 · outbound

This paper cites An investigation of generative replay in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning An investigation of generative replay in deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.089516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.264341Z digest=sha256:6db2e365d3fb0073f8ca0c8b79505e47c1c7b76eb11b0f62d404bf7de19c61ed

Observation f8429033-e01f-4244-9be0-55bba015ae5e · outbound

This paper cites Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:32:55.635890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.310272Z digest=sha256:29d156839082c3007b7ae17b27369633a929a03013930b6fbefe3fb1c8606086

Observation f931b24d-da63-42ea-8c8b-e07459ba2322 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.372728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.372728Z digest=sha256:575167ee885eadecd686f2717f321b8a40d2939a08bff6c38c0d142b93c97c65

Observation 7f149a22-56a5-47e0-b243-32fd713826ff · outbound

This paper cites CURL : Contrastive unsupervised representations for reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.080060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.422423Z digest=sha256:dba32f16994934ab4a58b5ada9ab54d63ab336c5f40166aed0b0fc72cc6d1f35

Observation 4eba7925-a8d6-498a-bb0c-2245d0a24203 · outbound

This paper cites HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.484896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.484896Z digest=sha256:94c3925cf867c6b849c6e129fb2a95a966217e90f3a944d78d3caa8959a370eb

Observation 9027ad0e-f8f2-4657-a7be-1ecfbd57f57a · outbound

This paper cites Continuous control with deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Continuous control with deep reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.581844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.581844Z digest=sha256:c5a14a2a8949e3003849840d437a24cc7a63aeab0b0a254c23dd4b771760531c

Observation 8ca1f07c-c3b9-4933-8c83-ec8274f9d30f · outbound

This paper cites Unlock the intermittent control ability of model free reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unlock the intermittent control ability of model free reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.069185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.627229Z digest=sha256:77b5f1127f5480e3929e806bf49e2a414af4efc218a1a69903aabf5f05ec1fa7

Observation 66a22057-2102-44d2-91ba-467c9ce178c7 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.060067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.676803Z digest=sha256:e8d8ad2d2fdfbe828a922071318fe094c762b841c9bf298c47a69fedb614f026

Observation 777ddf33-8dd3-45f6-841b-b27fa644f0c7 · outbound

This paper cites W., and Parker-Holder, J.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., and Parker-Holder, J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.049548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.718788Z digest=sha256:795ef90eb91a7049dc9f92fbd23ce52dd9a25ea81fbee85223cc759547542c19

Observation 5a3c5aca-0cec-4191-a7ba-ccd0c65ce5ea · outbound

This paper cites Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.760568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.760568Z digest=sha256:02df36d2ccd435933abcae16c40228cd1c8629af21141b4473eb2aecf656c153

Observation 914aae6e-5090-4352-802d-356c8f4d15ef · outbound

This paper cites Online reinforcement learning with uncertain episode lengths.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Online reinforcement learning with uncertain episode lengths

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.040234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.823709Z digest=sha256:796fa902fea21e482c61ae7f2cb618069a8d61645e35c4597acb3f1c6d56a41e

Observation 1bada11b-e190-455f-b112-7defdb9beb3d · outbound

This paper cites Tactical optimism and pessimism for deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Tactical optimism and pessimism for deep reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.029700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.877004Z digest=sha256:7e913767b734bddc934cefccdc37549ab2024e19db54e3b2c0913c78190eb5e1

Observation 0d1d9e27-5aed-4b82-a953-88d50f1a3aeb · outbound

This paper cites Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.919314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.919314Z digest=sha256:4848b1b56e28170a8b2f48d1d1a86f66c899bca2c60d17c008815ff7d522fff2

Observation 070c5cfe-e979-4beb-bcc9-a55dbd64ec20 · outbound

This paper cites The primacy bias in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning The primacy bias in deep reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.016950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.968328Z digest=sha256:b0c6fd092ca42d65b20ae5f983a6f92fb157a3af722acf3664e4f0c68ac01c76

Observation e8eb03d9-2105-45a2-bbaa-990e2397a778 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.017309Z digest=sha256:40e7e0f71c6d8db750c86f9508ae389a555800c9c99b597a74910f2a1a0905e9

Observation c0fc6932-682d-482a-bfc2-4920c266cc36 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.992729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.051990Z digest=sha256:35373b1145304925c30de2496408aac9e5afa6d2235dd42f7e0135fb4753209c

Observation 8668d5d8-ef85-4d7b-8892-2c64c2f76568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.099844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.099844Z digest=sha256:d6299854dde21d55361790583fba3aa8467f7b56dab6cc93bcf7ef3f634d6c47

Observation 65a0b996-78b1-4ed1-a3ea-e92db6cdf013 · outbound

This paper cites Reinforcement Learning with Dynamic Boltzmann Softmax Updates.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Reinforcement Learning with Dynamic Boltzmann Softmax Updates

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.160250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.160250Z digest=sha256:c6b5857d23367e609ad39e6676a8ec47346a0865406a4df0ea93296bb5097b50

Observation 9660b867-6b68-419b-88d0-fd77ae2cf9f2 · outbound

This paper cites Softmax deep double deterministic policy gradients.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Softmax deep double deterministic policy gradients

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.981193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.192637Z digest=sha256:bc4f794ddb376b524d4b1625ae2082531084c758bb7ca66fd9cef1d364152ff3

Observation fd8f489c-b566-440a-88c7-cf0253156be3 · outbound

This paper cites Regularized softmax deep multi-agent q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Regularized softmax deep multi-agent q-learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.970541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.230814Z digest=sha256:7c3146094e22bdc62dc2461c1589f5ab60a2f82b1fc6dd1d745d9431f09f7244

Observation 5c1641d6-a350-4942-84a0-4fc0f58aa313 · outbound

This paper cites Time limits in reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Time limits in reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.958785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.321621Z digest=sha256:da4c54776c16057a632da8081f2ae8e6b1ff35ee3a5929bce2ce5eff4fa17886

Observation a39b91ca-6935-48b4-a1b8-47d17ec83dd5 · outbound

This paper cites A., and Darrell, T.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning A., and Darrell, T

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.947216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.386837Z digest=sha256:7299200e5960b367a79d3d202e594b3d0d9a12d21b4eb87be16fbaedf85c7f07

Observation 2986c97b-8c8d-4f18-b235-2c43b089d4f4 · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.934747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.430420Z digest=sha256:856ffda39532c7a549a499c11b81d2a13fd6e5314200f0d63c06c57e15beb325

Observation 42e3395c-bc25-4354-8cb0-42c32b25495c · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.921337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.493835Z digest=sha256:82deab23417dde34655798ff77d7687eca1f1a711d32b02e8b1d2fd9aef474f8

Observation ac7e3916-194d-49ab-9420-f5302e4f0da4 · outbound

This paper cites Optimistic Exploration even with a Pessimistic Initialisation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimistic Exploration even with a Pessimistic Initialisation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.555362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.586211Z digest=sha256:7c8af2bcca042d65cddfe192be20b5fb6d9ebce3d030a8deac51f7f4a1459856

Observation b7cee4a3-ece9-4414-85ac-3e7ec8f782a7 · outbound

This paper cites Prioritized Experience Replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.629886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.629886Z digest=sha256:4d9b0b92b425ff49309ce92c4c1f6bd8081476e4d2a803e56bfe47f5b4b61b10

Observation 0e458a88-e4c5-4099-9924-1a5abfd97568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.908081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.697565Z digest=sha256:4a3863b5d4b38213755dfd22d47fcdf9dbf79ac5e9b9fd4e1b9672c967c004b8

Observation e9e3441b-e682-48b6-9f1b-4ea00bcb32b4 · outbound

This paper cites S., and Evci, U.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning S., and Evci, U

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.837429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.837429Z digest=sha256:c75c63cb41622e815e38451ef04ab55f842bed6767c0ba3d97f0f0d1e12d0a13

Observation 6a35abc8-8d00-4b36-893e-a38fb2906a11 · outbound

This paper cites Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.490510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.902691Z digest=sha256:8c7f6709edcfa379e09e0c0059342b1954ccf648fa0c59ac21848fe80da6e2df

Observation f4ab1462-9de7-4f2a-8679-bc0723435502 · outbound

This paper cites Revisiting the softmax bellman operator: New benefits and new perspective.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting the softmax bellman operator: New benefits and new perspective

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.886519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.972733Z digest=sha256:36e9f6a34e0cb450c0df3094a2e806f4eafbba0a205743dc657f1b954b9b5301

Observation d90db58a-4793-4d0b-9c9c-9099759ce176 · outbound

This paper cites Prioritizing samples in reinforcement learning with reducible loss.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritizing samples in reinforcement learning with reducible loss

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.876288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.084002Z digest=sha256:fcaa3fd44de791e117a96b7c55a1b2caf218c9defcbc4a6ea87937975990b760

Observation e6907165-b461-4f4a-ad31-8153b8e506b7 · outbound

This paper cites Safe Exploration by Solving Early Terminated MDP.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Safe Exploration by Solving Early Terminated MDP

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.325475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.197634Z digest=sha256:1ffde16e38c8f63e85f359ae0274a2b98665c7b7aeb5d25d056e4eecdbe55910

Observation 35d61335-ae2b-479d-abab-982bc9a399e0 · outbound

This paper cites sunk costs.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning sunk costs

Reference 56

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.587105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.278813Z digest=sha256:895bc2e7096abde8e0cb4b10e3e66b68fbb17ecbc9a75b37bddb0d6d7dc4529d

Observation 364092c5-6388-4b31-afdc-14272e675950 · outbound

This paper cites DeepMind Control Suite.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DeepMind Control Suite

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.351761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.351761Z digest=sha256:8a3dda25163f7e5bc6a56a773d4b8f5161eea595414f486116c2c9a044bc9b2a

Observation 6e83b4c8-5af9-4201-9939-05fa456612a6 · outbound

This paper cites Loss Functions and Metrics in Deep Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Loss Functions and Metrics in Deep Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.432291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.432291Z digest=sha256:6cfb09ce011d38b91153c28db9cd7b3bb7313ab6abd4402909625dc99ed2451e

Observation 388bbcb7-8741-43ce-a4a1-fe19799cd702 · outbound

This paper cites H., Meyers, E.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning H., Meyers, E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.866665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.623704Z digest=sha256:ce80f45b2d6168df618f7eb2b934904a7e13aef8116590267176bed96402ccb7

Observation 31640fed-96f4-4f9b-9cd6-58fb38c952ce · outbound

This paper cites Deep reinforcement learning with double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Deep reinforcement learning with double q-learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.857507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.697176Z digest=sha256:6341090560374fb7944e60aeb4aa8899d4edae2312d8fe8360902b7936b05a68

Observation 58976586-1628-4cf5-a242-d26288add4d0 · outbound

This paper cites and Drake Jr, F.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning and Drake Jr, F

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.847097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.713465Z digest=sha256:2fb1efd779dd9509a0919e20eec45e8e51341cf2ac80b868799b2e2a37102def

Observation cc25d49c-21eb-4831-96f8-e5b9a1915127 · outbound

This paper cites Optimism in Reinforcement Learning with Generalized Linear Function Approximation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimism in Reinforcement Learning with Generalized Linear Function Approximation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.761065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.761065Z digest=sha256:97978e7c6c1ae59ca996dc2f6049d53b85e4b6df4e02e3eee913a98a66c768b1

Observation 0f10747d-6b07-4af0-b75b-552092ec3032 · outbound

This paper cites DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.901414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.901414Z digest=sha256:9ccec6163dcd0919d094b6da581dc726dcaf5d91f0339968f807813811fb537f

Observation c727a42e-9e6c-4af3-97cf-7fda9d49554c · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.985431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.985431Z digest=sha256:b233d1039e447fa0f4a768170a9d71e037fe7f56d17971ad544321ddb93fc139

Observation a72e8898-159d-41b9-a619-b35c1687e2dc · outbound

This paper cites Sample Efficient Deep Reinforcement Learning via Local Planning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sample Efficient Deep Reinforcement Learning via Local Planning

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.136783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.102945Z digest=sha256:0e4383fab8366c602211af9da6d6db424733be1e94abda0b2f0523e6596fbcc6

Observation 0475043e-986d-4206-bde3-a48ea046d7fb · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Scaling Robot Learning with Semantically Imagined Experience

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.190932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.190932Z digest=sha256:f0648d9e845a4379ec0e722c48e51a851276cfeefa9f31fed96a8ec541948754

Observation 0b575bae-4391-42df-9f0a-3a9a92463523 · outbound

This paper cites Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.835270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.287162Z digest=sha256:b4489e6e3467803a4360bc9870deae26dffe9d65488c1c61ac92171f86bb70dc

Observation ea620ca8-f979-4a01-ac43-e0762f4be893 · outbound

This paper cites write newline.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning write newline

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.400821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.400821Z digest=sha256:6339ac81689d2a1a7344a749634450cf3a7e8a84a5c4898e7b8cb3c8d238e3f3

Pith citing papers

No inbound Pith citation observations are available.