Pith. sign in

Paper Citation Record · LEDGER

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2506.13672.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13672 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:32:54.400821Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact9
  • verified fuzzy21
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 088c0f74-fe5f-4416-898a-0d8fa0a6237d · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.191659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.034330Z digest=sha256:96c29cbe9c11f4a27f23147ec86f956399ec0987a86abae385eb280203963d2c

Observation d8740eef-1ef0-4a05-a088-89479bf979f8 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.182111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.070341Z digest=sha256:0eefcbfac6b4c856e83959121454eee803157076f0936a73d9387fb1144f9217

Observation ba96e2ff-57cd-4c4e-be56-0657917e6209 · outbound

This paper cites OpenAI Gym.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.109869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.109869Z digest=sha256:872f51d2a66e1df59e10acfae7a1493a40628df2f3c79a058f2588528fc81a2e

Observation c78b0f73-f7eb-439d-9fa3-82d439d71346 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.172292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.173653Z digest=sha256:4e9318ba14d343d31c3ce053f3dadac8f24f4fbd39ae489567f778fc3f80725d

Observation 94e30e18-13d6-4818-b1e4-163f389d09a6 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.161746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.232671Z digest=sha256:0aff2e2a2f07563ba81b772e19c091be2e373d06f4df3289d58fa932f05fc554

Observation a084e66b-bf25-4ca5-b7ef-e03e92b9180e · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.151864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.284609Z digest=sha256:15ceb5f0023be558781d634d0e9be610eb0a985ae2641382cae82a0648a74eaa

Observation ac4175c7-aa18-49a8-8509-72ac5a37170c · outbound

This paper cites Stabilizing Off-Policy Deep Reinforcement Learning from Pixels.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Stabilizing Off-Policy Deep Reinforcement Learning from Pixels

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.368740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.368740Z digest=sha256:3e8f6343b942c61e455741603233ddce54280d705a16a870d8d5125d443c43d7

Observation b58b6687-ec6e-4d18-b8aa-6fc75da5b71c · outbound

This paper cites Randomized Ensembled Double Q-Learning: Learning Fast Without a Model.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Randomized Ensembled Double Q-Learning: Learning Fast Without a Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.412668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.412668Z digest=sha256:9e7250a0d1af07a0d12a317bd6106ddcd353cdb9fbeacef6b4471636ccb5d3c0

Observation 47f52821-2c7f-4b10-a141-14f7bd75a73f · outbound

This paper cites Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.773730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.489700Z digest=sha256:1e3c93724a6be796311b32dc66a646855c813bcf687287c8a940f055fbc1eb50

Observation 134041a6-9e97-48bb-8d6f-ba2341c7ff5f · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.925914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.552848Z digest=sha256:0a3652f41f6fcc4b68784bb5ea1ca87ee3a8a01e1a891962ed06778fb358886a

Observation a8db4646-1acd-404e-87ea-702597f337cb · outbound

This paper cites G., and Courville, A.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning G., and Courville, A

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.592830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.592830Z digest=sha256:e57124b1f85f32fb611e3b5a8cdfa8155f99b86026500abfdf6ef1edfc050508

Observation 74eb255d-4dae-4ba4-9b9c-1872bf0fcfcd · outbound

This paper cites W., Subramanian, J., and Ghassemi, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., Subramanian, J., and Ghassemi, M

Reference 12

Resolution
verified exact
raw_fallback, observed 2026-08-07T00:32:55.758573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.628887Z digest=sha256:2b0ddba862e69ba27310d15bd9d431a579f88032372636387ba9dadedf7a39e6

Observation 67b3ad1d-71b8-4f47-973f-7ce034d107c7 · outbound

This paper cites Revisiting fundamentals of experience replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting fundamentals of experience replay

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.703013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.703013Z digest=sha256:3dc8020efa20a99a6cd8e5b59cd072d9cf79f4fe563ae24b7fc61d234abe3def

Observation b23c158e-46a9-4aa9-b942-a735050d9861 · outbound

This paper cites Addressing function approximation error in actor-critic methods.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Addressing function approximation error in actor-critic methods

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.764866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.764866Z digest=sha256:a5d006b87083265f5d2340176b7f0972a797e82a4ce531b9e7856978809dc27a

Observation c08bb5d9-d454-446f-b43f-21835d58dad6 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:50.832637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:50.832637Z digest=sha256:2512fd26703cf072f0f8afb394b45809332920bd4f9001b1a37c3f2de7153e80

Observation 61054b92-983b-4b0b-831e-0dc697467c82 · outbound

This paper cites Sunk-cost fallacy and cognitive ability in individual decision-making.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sunk-cost fallacy and cognitive ability in individual decision-making

Reference 16

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.706719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.905906Z digest=sha256:04c1ef319d43d58b748b1ef1574302ab5e2786abf6b4400441abfcf98c82e2e0

Observation 234a0dae-0fd4-42b5-a57c-451e32a7c932 · outbound

This paper cites R., Millman, K.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning R., Millman, K

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.116543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:50.926428Z digest=sha256:190c41b5080b615df230c917f9de7b8ae296e9d701630065e3d5246bd3a0a94d

Observation cf7f6777-c012-435d-ab6e-bb8e2237927d · outbound

This paper cites Double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Double q-learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.107696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.025987Z digest=sha256:75110df12768abb94560bee341a44b8d23b380841328b489f82d45bb96b04227

Observation 962346cb-77aa-4e31-a6e4-c439b653c794 · outbound

This paper cites Dropout Q-Functions for Doubly Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Dropout Q-Functions for Doubly Efficient Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.089069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.089069Z digest=sha256:ecb16aed1eb49833d034909b970eb3a640b8e57cc0f2289942d2f48f99ec41e8

Observation 30c170d8-8f14-4a6c-8741-82a3045d317f · outbound

This paper cites Planning Goals for Exploration.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Planning Goals for Exploration

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.133496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.133496Z digest=sha256:768f37af94baf0b3fbc68080f40f0f652342495d63df97fd16f389c884b301bf

Observation 0b6767b6-f2ed-4dfd-8933-e080872a10b3 · outbound

This paper cites Enhanced Experience Replay Generation for Efficient Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Enhanced Experience Replay Generation for Efficient Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.201826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.201826Z digest=sha256:507316a3bd1b47f04c07ede61624f27a407667aef42115ade777c1a4f251d31f

Observation a7febbd2-a0e6-47f6-8055-a046938b0574 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.098144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.236555Z digest=sha256:d45bb0e371f239b870c17df98af41c2fb20baa262189af4bad6b929f58795343

Observation a28efe13-186d-491b-926d-7db7514d8bb1 · outbound

This paper cites An investigation of generative replay in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning An investigation of generative replay in deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.089516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.264341Z digest=sha256:958d208bee59a2556e973e634d6f11a985c94c4ce5fdceb545cf41cc5e5c192b

Observation f8429033-e01f-4244-9be0-55bba015ae5e · outbound

This paper cites Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Risk Sensitive Dead-end Identification in Safety-Critical Offline Reinforcement Learning

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:32:55.635890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.310272Z digest=sha256:ea6117f1374c04a1af7ac1f33385a0dee445af77c2c867f0719788c0ecf519f9

Observation f931b24d-da63-42ea-8c8b-e07459ba2322 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.372728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.372728Z digest=sha256:c4d8dd983697030ec57d353a679e5d4901077f4cce4d67b43b340e95ec01e420

Observation 7f149a22-56a5-47e0-b243-32fd713826ff · outbound

This paper cites CURL : Contrastive unsupervised representations for reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.080060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.422423Z digest=sha256:1784e424b1b325cbd3dd4dca6337839a968a5051b6a60f5309b9623692f747e0

Observation 4eba7925-a8d6-498a-bb0c-2245d0a24203 · outbound

This paper cites HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning HyAR: Addressing Discrete-Continuous Action Reinforcement Learning via Hybrid Action Representation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.484896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.484896Z digest=sha256:94c3925cf867c6b849c6e129fb2a95a966217e90f3a944d78d3caa8959a370eb

Observation 9027ad0e-f8f2-4657-a7be-1ecfbd57f57a · outbound

This paper cites Continuous control with deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Continuous control with deep reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.581844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.581844Z digest=sha256:c5a14a2a8949e3003849840d437a24cc7a63aeab0b0a254c23dd4b771760531c

Observation 8ca1f07c-c3b9-4933-8c83-ec8274f9d30f · outbound

This paper cites Unlock the intermittent control ability of model free reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unlock the intermittent control ability of model free reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.069185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.627229Z digest=sha256:ffe4d69f85f5a30ccfbe18376dbc336abed87d49ec7ff49292c5c63091926a72

Observation 66a22057-2102-44d2-91ba-467c9ce178c7 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.060067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.676803Z digest=sha256:0d0bca7128925d805959a8c6fcdd03b6d69e1bfc588d0556f4e24cb23a8753ad

Observation 777ddf33-8dd3-45f6-841b-b27fa644f0c7 · outbound

This paper cites W., and Parker-Holder, J.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning W., and Parker-Holder, J

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.049548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.718788Z digest=sha256:bac7c59b1425c00205d4740faf51c300ff1e0a7f7dca257a48f641b4d38ab1e5

Observation 5a3c5aca-0cec-4191-a7ba-ccd0c65ce5ea · outbound

This paper cites Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.760568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.760568Z digest=sha256:02df36d2ccd435933abcae16c40228cd1c8629af21141b4473eb2aecf656c153

Observation 914aae6e-5090-4352-802d-356c8f4d15ef · outbound

This paper cites Online reinforcement learning with uncertain episode lengths.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Online reinforcement learning with uncertain episode lengths

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.040234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.823709Z digest=sha256:97acf1094251b7847be563a5191356d275eda179bc3ec322c66a238dcbf34bc5

Observation 1bada11b-e190-455f-b112-7defdb9beb3d · outbound

This paper cites Tactical optimism and pessimism for deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Tactical optimism and pessimism for deep reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.029700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.877004Z digest=sha256:5ffbc6910fa2d29fbe807d474f295a2f27333f5fa1ca1147dd0ee49682fdc480

Observation 0d1d9e27-5aed-4b82-a953-88d50f1a3aeb · outbound

This paper cites Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Bigger, Regularized, Optimistic: scaling for compute and sample-efficient continuous control

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:51.919314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:51.919314Z digest=sha256:da96e2fb2dbefc0c8e01027773fe23a51defa71140ca47227110d520888b2463

Observation 070c5cfe-e979-4beb-bcc9-a55dbd64ec20 · outbound

This paper cites The primacy bias in deep reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning The primacy bias in deep reinforcement learning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:56.016950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:51.968328Z digest=sha256:dc91c0ca75cc3b2c7e545081c12e81d109bc3f0eedc33897e7d45f14c9ca5bfb

Observation e8eb03d9-2105-45a2-bbaa-990e2397a778 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.017309Z digest=sha256:cf0afeec7362c494aa47a1f171dd09eb8d501e17f129bc9e1abdd4f798423245

Observation c0fc6932-682d-482a-bfc2-4920c266cc36 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.992729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.051990Z digest=sha256:d3c7b64efe3df1e31bd12a32aa16330a04e9f21eca21505f8b2aeaeea8bce73e

Observation 8668d5d8-ef85-4d7b-8892-2c64c2f76568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.099844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.099844Z digest=sha256:d6299854dde21d55361790583fba3aa8467f7b56dab6cc93bcf7ef3f634d6c47

Observation 65a0b996-78b1-4ed1-a3ea-e92db6cdf013 · outbound

This paper cites Reinforcement Learning with Dynamic Boltzmann Softmax Updates.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Reinforcement Learning with Dynamic Boltzmann Softmax Updates

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.160250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.160250Z digest=sha256:c6b5857d23367e609ad39e6676a8ec47346a0865406a4df0ea93296bb5097b50

Observation 9660b867-6b68-419b-88d0-fd77ae2cf9f2 · outbound

This paper cites Softmax deep double deterministic policy gradients.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Softmax deep double deterministic policy gradients

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.981193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.192637Z digest=sha256:0d12dd4b08cfe5793e1505c5af494d9ae52b93a9c46ed2d66b6ac4894beebfd9

Observation fd8f489c-b566-440a-88c7-cf0253156be3 · outbound

This paper cites Regularized softmax deep multi-agent q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Regularized softmax deep multi-agent q-learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.970541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.230814Z digest=sha256:06254cd0a37ac1b04f4cb16cbc9645ef9081bbf1fd0c241e25522ff49df63e89

Observation 5c1641d6-a350-4942-84a0-4fc0f58aa313 · outbound

This paper cites Time limits in reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Time limits in reinforcement learning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.958785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.321621Z digest=sha256:8951b58f8c0108399879096ce0da4d8a7d7b2f84282e57a1cf355311fe0aa05f

Observation a39b91ca-6935-48b4-a1b8-47d17ec83dd5 · outbound

This paper cites A., and Darrell, T.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning A., and Darrell, T

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.947216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.386837Z digest=sha256:6387f30c8d51c3e5a6fbd1ecd1d3813857d85a11aaaf667cbd2ecd60c3fd8853

Observation 2986c97b-8c8d-4f18-b235-2c43b089d4f4 · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.934747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.430420Z digest=sha256:c109bb0f031a78ab9fc311f027068740572857ed316c1e14a471b99cd8151ee9

Observation 42e3395c-bc25-4354-8cb0-42c32b25495c · outbound

This paper cites M., and Restelli, M.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning M., and Restelli, M

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.921337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.493835Z digest=sha256:ddc54e9996c29529f77a4eef87c79104212f9b601f5a457e3199ade415dca14e

Observation ac7e3916-194d-49ab-9420-f5302e4f0da4 · outbound

This paper cites Optimistic Exploration even with a Pessimistic Initialisation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimistic Exploration even with a Pessimistic Initialisation

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.555362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.586211Z digest=sha256:b006838829d5fd5a57958043fb99a57df10f320e177556997a52f4cd1307416b

Observation b7cee4a3-ece9-4414-85ac-3e7ec8f782a7 · outbound

This paper cites Prioritized Experience Replay.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritized Experience Replay

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.629886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.629886Z digest=sha256:4d9b0b92b425ff49309ce92c4c1f6bd8081476e4d2a803e56bfe47f5b4b61b10

Observation 0e458a88-e4c5-4099-9924-1a5abfd97568 · outbound

This paper cites an unresolved cited work.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:32:55.908081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.697565Z digest=sha256:17e27c3164b505047db9acbef78fc016c661ac951c9f6db430b839942c738278

Observation e9e3441b-e682-48b6-9f1b-4ea00bcb32b4 · outbound

This paper cites S., and Evci, U.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning S., and Evci, U

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:52.837429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:52.837429Z digest=sha256:c75c63cb41622e815e38451ef04ab55f842bed6767c0ba3d97f0f0d1e12d0a13

Observation 6a35abc8-8d00-4b36-893e-a38fb2906a11 · outbound

This paper cites Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.490510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.902691Z digest=sha256:f6b261210923861d2c2ca66f0b5317ab3f28f364997263b95d0d7a661d51651b

Observation f4ab1462-9de7-4f2a-8679-bc0723435502 · outbound

This paper cites Revisiting the softmax bellman operator: New benefits and new perspective.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Revisiting the softmax bellman operator: New benefits and new perspective

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.886519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:52.972733Z digest=sha256:5f00fd793744c1b59065c91f51f0b26e7b8f3b847d4345031e7fe759b12b2451

Observation d90db58a-4793-4d0b-9c9c-9099759ce176 · outbound

This paper cites Prioritizing samples in reinforcement learning with reducible loss.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Prioritizing samples in reinforcement learning with reducible loss

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.876288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.084002Z digest=sha256:ead083cd25b67af20bb1f919174ee07138ae5422fb771a8093cd3e24b058363b

Observation e6907165-b461-4f4a-ad31-8153b8e506b7 · outbound

This paper cites Safe Exploration by Solving Early Terminated MDP.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Safe Exploration by Solving Early Terminated MDP

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.325475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.197634Z digest=sha256:39ab13a00e0a063c3aa6d247c640ce2fe29aadbc8d83f3db0a44a54bc500e2fd

Observation 35d61335-ae2b-479d-abab-982bc9a399e0 · outbound

This paper cites sunk costs.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning sunk costs

Reference 56

Resolution
verified exact
doi, observed 2026-08-07T00:32:54.587105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.278813Z digest=sha256:adbf4265352f4e5f8e7c882f1f1467a2d7f34bebd019525fc27643bb7cfcc52f

Observation 364092c5-6388-4b31-afdc-14272e675950 · outbound

This paper cites DeepMind Control Suite.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DeepMind Control Suite

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.351761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.351761Z digest=sha256:8a3dda25163f7e5bc6a56a773d4b8f5161eea595414f486116c2c9a044bc9b2a

Observation 6e83b4c8-5af9-4201-9939-05fa456612a6 · outbound

This paper cites Loss Functions and Metrics in Deep Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Loss Functions and Metrics in Deep Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.432291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.432291Z digest=sha256:6cfb09ce011d38b91153c28db9cd7b3bb7313ab6abd4402909625dc99ed2451e

Observation 388bbcb7-8741-43ce-a4a1-fe19799cd702 · outbound

This paper cites H., Meyers, E.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning H., Meyers, E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.866665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.623704Z digest=sha256:a912a756846598129f196043324df7313339fba7b3bf3266e423cb4ef4b8041e

Observation 31640fed-96f4-4f9b-9cd6-58fb38c952ce · outbound

This paper cites Deep reinforcement learning with double q-learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Deep reinforcement learning with double q-learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.857507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.697176Z digest=sha256:61dc2a42116ba29b2e3797c0363d1d5102d5e0c9d6cbf3fa7471b99982aee41f

Observation 58976586-1628-4cf5-a242-d26288add4d0 · outbound

This paper cites and Drake Jr, F.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning and Drake Jr, F

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.847097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:53.713465Z digest=sha256:1d1d9df3e75fa9f0a0862a059df13a7ad931b0932a278791a8b42a67d0bce6b8

Observation cc25d49c-21eb-4831-96f8-e5b9a1915127 · outbound

This paper cites Optimism in Reinforcement Learning with Generalized Linear Function Approximation.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Optimism in Reinforcement Learning with Generalized Linear Function Approximation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.761065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.761065Z digest=sha256:97978e7c6c1ae59ca996dc2f6049d53b85e4b6df4e02e3eee913a98a66c768b1

Observation 0f10747d-6b07-4af0-b75b-552092ec3032 · outbound

This paper cites DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning DrM: Mastering Visual Reinforcement Learning through Dormant Ratio Minimization

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.901414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.901414Z digest=sha256:9ccec6163dcd0919d094b6da581dc726dcaf5d91f0339968f807813811fb537f

Observation c727a42e-9e6c-4af3-97cf-7fda9d49554c · outbound

This paper cites Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Mastering Visual Continuous Control: Improved Data-Augmented Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:53.985431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:53.985431Z digest=sha256:b233d1039e447fa0f4a768170a9d71e037fe7f56d17971ad544321ddb93fc139

Observation a72e8898-159d-41b9-a619-b35c1687e2dc · outbound

This paper cites Sample Efficient Deep Reinforcement Learning via Local Planning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Sample Efficient Deep Reinforcement Learning via Local Planning

Reference 65

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:32:55.136783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.102945Z digest=sha256:b288e212f46704548b0dd348f4bfa0bbbfb0cad9a66b5f702e5907e937ec814f

Observation 0475043e-986d-4206-bde3-a48ea046d7fb · outbound

This paper cites Scaling Robot Learning with Semantically Imagined Experience.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Scaling Robot Learning with Semantically Imagined Experience

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.190932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.190932Z digest=sha256:f0648d9e845a4379ec0e722c48e51a851276cfeefa9f31fed96a8ec541948754

Observation 0b575bae-4391-42df-9f0a-3a9a92463523 · outbound

This paper cites Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning Taco: Temporal latent action-driven contrastive loss for visual reinforcement learning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:32:55.835270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-07T00:32:54.287162Z digest=sha256:b38ce968cda2f7c528aeb56cbe2554fdae0e14eab8c3e9b1f6f1c8e971f8876a

Observation ea620ca8-f979-4a01-ac43-e0762f4be893 · outbound

This paper cites write newline.

The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep Reinforcement Learning write newline

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T00:32:54.400821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:32:54.400821Z digest=sha256:6339ac81689d2a1a7344a749634450cf3a7e8a84a5c4898e7b8cb3c8d238e3f3

Pith citing papers

No inbound Pith citation observations are available.