Pith. sign in

Paper Citation Record · LEDGER

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 2 inbound Pith citation observations for arXiv:2508.04216.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.04216 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:49:58.910725Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:22:06.635622Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T12:23:24.240761Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 53475a56-da37-4582-bb18-dfc7b62a9ad8 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.195733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.195733Z digest=sha256:5ca178803c71961c2469f71e74704ef1966223679c5c2aedc27f96e6ee60f2b5

Observation a21285f0-4606-462a-8b62-e26e1c89cb44 · outbound

This paper cites write newline.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.287688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.287688Z digest=sha256:3146d3bc2b60f70fdceedd51b98eaf5104fd96fbbaa1c303459988f616f754d8

Observation 7953bb5a-bb0d-4a5f-bbc0-e78ea0e11043 · outbound

This paper cites Concrete Problems in AI Safety.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Concrete Problems in AI Safety

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.417778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.417778Z digest=sha256:b1920119495c73c8d23bbc4902f38ad7b246c23bb25fd247ce1327807366f34c

Observation 426f3b28-a921-476d-88da-4a3fe872c68c · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:02.140984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:55.552626Z digest=sha256:4e129677833a9da73de6ebbaf7d52e3a0153b225b0d81108ee6209a6410c9754

Observation 870055df-2744-4273-8716-6b2db294e08f · outbound

This paper cites E.; Hume, T.; Carter, S.; Henighan, T.; and Olah, C.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction E.; Hume, T.; Carter, S.; Henighan, T.; and Olah, C

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.665028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.665028Z digest=sha256:64616eeda4efd1e737b771def42682f7975f6f79ae109b4b0a5ace78ea89548a

Observation d56d4c56-8dea-4b89-a9fe-5ff1ff83ac11 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.743365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.743365Z digest=sha256:aa8b361dfbde4497f4dca54cd84132c1023823b126a5bd08b2e97be1daf76e87

Observation 7532f3ab-76f8-46a2-9748-bac3186db987 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.811698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.811698Z digest=sha256:d44989c0bd56f84e756c277f439f22ace1d762344fd36eb712f8092404e29277

Observation f4fed5e4-9b22-432b-9197-85b9a818c276 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.891955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.891955Z digest=sha256:7c27a8be61f1568ef620f346726205f342fd1acc50b24864ccbe4bd98fe4bd98

Observation 48843abe-3a9f-443d-aca0-0ce19980c0d0 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:55.975299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:55.975299Z digest=sha256:55c1365cc20409b4776c071c5f0b58478ea21b43c6078e88be7363c07a5ca893

Observation c175fe98-219e-43b5-be7c-dcaddc4ccf1d · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:01.974682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:56.047481Z digest=sha256:0b8b46d4107b12b84e1ae0b457b6a78f904037c974122bd015a34c688341ab75

Observation 6b720003-3c15-4315-87ef-c01af2f16297 · outbound

This paper cites The Llama 3 Herd of Models.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.114202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.114202Z digest=sha256:e864dba7e1ce58627a1fc1aa33072963e0e0ca5a6872ee47f3c583624da5cd05

Observation ff348be9-27de-427a-8ce6-a0f67c93a558 · outbound

This paper cites Inverse Reward Design.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Inverse Reward Design

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.182077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.182077Z digest=sha256:40e482ed0ca227477e399703b980bb278d24ef2f5c0d7a00e37035b31303b162

Observation 993c3d33-30a8-4051-9fcd-da0bce7e8267 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:01.825837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:56.235762Z digest=sha256:d3aa2715fc3d3aebec586ce6cc3fce56eb488b0b2111ed7b74343e69bed68696

Observation 02c988b7-9dd4-47ae-bba3-607f415acebc · outbound

This paper cites Let's Verify Step by Step.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Let's Verify Step by Step

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.306521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.306521Z digest=sha256:161e1bc3598052dbd2b7d71590ac151a1d40f779751423284e585bf3e620621c

Observation 9c1a11d3-de30-4c11-9006-604f787b15e8 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.372971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.372971Z digest=sha256:0afee2b546b3a7442efaf228a12cbaea95c12410c85b0da817034c170078bb55

Observation 151a3599-716c-4f98-ad09-f762cebc36b2 · outbound

This paper cites RRM: Robust Reward Model Training Mitigates Reward Hacking.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction RRM: Robust Reward Model Training Mitigates Reward Hacking

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.466791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.466791Z digest=sha256:20ec55ff3d1df1dbc49a7cc6eee13efa09d9ed20ec4a3e040e09d35a7655b7bb

Observation d2786640-1018-43ad-8f4e-99a900c9a45c · outbound

This paper cites P.; Hermann, K.; Welleck, S.; Yazdanbakhsh, A.; and Clark, P.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction P.; Hermann, K.; Welleck, S.; Yazdanbakhsh, A.; and Clark, P

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:50:01.604360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:56.543973Z digest=sha256:cacd282d1e5739d8cbe37c2bade4db01d68a3564b2f7404e5b56de568ae94fed

Observation 41967675-9e22-4ff0-9d31-615883143be6 · outbound

This paper cites Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.667273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.667273Z digest=sha256:14d74a81666fda0316fb278a58ba64b24ee0cb1e99600ebb107ae7014cc98fcc

Observation f01d6fc3-fd5d-46be-9266-4d1dc61a95e1 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:01.465055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:56.757497Z digest=sha256:bc928aa098b4c741791b071325311c2a4bd4c95f8698d81275fca809d2a0a707

Observation 5fc8f8d2-3bc5-4065-acc1-1db253b092fd · outbound

This paper cites Feedback Loops With Language Models Drive In-Context Reward Hacking.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Feedback Loops With Language Models Drive In-Context Reward Hacking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.840261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.840261Z digest=sha256:e533f53d6c6d9c00e0ba75e4842b18f7c96301d8d38b1ed2e363d915b87828e3

Observation a8c35825-02bd-434b-9fd9-eacef2c06611 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:56.900164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:56.900164Z digest=sha256:f13d060292f45e2c5ebb183adb45256d7ef1598b7087c9a5e5d1ec98e0ad6782

Observation b599c292-0fe5-4857-8b69-41716d6dfa30 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:01.243287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:56.989041Z digest=sha256:460cc87c74abfd10af12bc23f6f2749c903523b5ff8e0e2fe22af496d4175d73

Observation 11feddce-4347-4d37-8522-ca4ea2a71324 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:01.071894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:57.073629Z digest=sha256:6ca7cce85c6e5bda61fb6bfc0b7ef8ac831ea926fa3eac9d5d7437a34a1500b3

Observation c81a857b-41ce-4035-aa10-4432d0b99b1f · outbound

This paper cites Self-critiquing models for assisting human evaluators.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Self-critiquing models for assisting human evaluators

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.156128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.156128Z digest=sha256:a11d42db75222822af846ea31375f0e3f9a7cad8121879eceb2d04dca58e4450

Observation 1aef55fa-a437-473b-bcd4-c93e13044071 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:00.812569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:57.244923Z digest=sha256:9d5c6ea44628cee57f03ed45d4b13386798ff5741874dfea6807e988cab5cc28

Observation 3abde802-218f-44d7-ba48-e00c9527e919 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:00.557972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:57.313769Z digest=sha256:f22406ca4be17e79ad68d8ffadc846bd6e58b82f5d24d9de1083217f8e2cc485

Observation d45b03e4-3e88-475d-a03f-6465fd3d4c36 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.391495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.391495Z digest=sha256:d07b0f170c9233d686af85d9ada3c42bc1f909d01ca4ef45164bd047a6ebd267

Observation 8ca4aa02-c030-4f7d-9025-fd91360e5b06 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.500199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.500199Z digest=sha256:d4d0fc5c0d24f303568343cd26278403d3234ad73149cdf94ab39d04e73f192e

Observation c4bf3754-bb08-44c9-8096-667569423794 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.585878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.585878Z digest=sha256:d622d38ea325b68642ae1668147bc294c99a1ba300648351ccf6e7e9f6156369

Observation f8048e73-95ee-4f11-8082-d620ed4e4900 · outbound

This paper cites V.; Lee, J.; Xu, K.; and Kumar, A.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction V.; Lee, J.; Xu, K.; and Kumar, A

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:50:00.332444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:57.685133Z digest=sha256:b32c26c2073cda141fc458ec83f3d4d6397912b101ab81f7e2d4a8279af5c9ec

Observation 94254e3b-f5aa-4a07-87f9-b90e7bc10d02 · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:50:00.119426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:57.794179Z digest=sha256:d49175c1a262fae8c3f043bdff4e2c59e93be867e45f686769f075ecbb7ad7f5

Observation 06cff769-dda9-49ac-877f-d6f01645f5de · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.868476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.868476Z digest=sha256:70bcacc8e2c6e98d2b1c406fb561955bc1ce880e3b5219be6980290930fb854f

Observation ea48494e-966a-4450-9252-0561cecf8267 · outbound

This paper cites L.; McDougall, C.; MacDiarmid, M.; Freeman, C.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction L.; McDougall, C.; MacDiarmid, M.; Freeman, C

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:57.951694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:57.951694Z digest=sha256:858d77a423a7a943520e9a9cf7e78a80c3a9cde338331fcff9468bef62c3066d

Observation 3fa39720-4336-4778-996c-69b162fd43dc · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Solving math word problems with process- and outcome-based feedback

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.030203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.030203Z digest=sha256:fb6a60e95e676db6492e047cf242c4a54ca3e054d713ac5bc6cfc83dcc5193cc

Observation 20e4ba43-a24d-4141-b814-b64088f259ae · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.128960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.128960Z digest=sha256:a0e57b191f90222fb4642a28f6c3f4ff16e44f60e9bc8146f7aa18f1b648a1a4

Observation 5210ed7d-8a0c-4e7d-8e2e-8a4285517db3 · outbound

This paper cites V.; and Zhou, D.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction V.; and Zhou, D

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:49:59.909412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:58.223409Z digest=sha256:c248cbd29acf2d2ee8087773e7876f380be9840926cb94953cb840a5af75aa0a

Observation bac4398b-b0d7-4bdf-9178-1545702fd1ef · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:49:59.733765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:58.324287Z digest=sha256:dbc4a96a70d5783bdcdcf02f41e3826f77778662a23fb51d323afccc91f991bb

Observation a68159bb-9a83-431f-834f-94b0921806ab · outbound

This paper cites an unresolved cited work.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:49:59.525941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:58.417699Z digest=sha256:a1483db63673fe83c3971bcac6d379585a75b934a750bfb436fba23d7b93c322

Observation 99f1c65a-14e5-4d15-8569-c13c054177e4 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.564042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.564042Z digest=sha256:aea2455d9682df06cf82ed13c6e93f377480ed23ba13d1493c29db93c82f1ab9

Observation 941c3b9b-e322-4766-ac12-5d0c6c6af4b9 · outbound

This paper cites L.; Cao, Y.; and Narasimhan, K.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction L.; Cao, Y.; and Narasimhan, K

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:49:59.304045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T00:49:58.645366Z digest=sha256:cf863314a4750c4056e600619eee26b598d609a74c2bc0c19658ada4ad4cc36f

Observation 87d5d540-9d51-4cb6-85c3-5b5292754f9a · outbound

This paper cites Automatic Chain of Thought Prompting in Large Language Models.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction Automatic Chain of Thought Prompting in Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.718193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.718193Z digest=sha256:a766d1d0fe885cbe0fcb5a2e6d9d3a8e3fa1f81475c60d5fcb5df42766783f09

Observation 7bb593af-902e-48f0-b2d3-c34250c53f0b · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.821187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.821187Z digest=sha256:2a32184ee886677eaa8886f2f1016900260f37802be907e419f6fdec158f9549

Observation 23122d00-eb26-4d5e-92d0-df77482177dd · outbound

This paper cites A Survey of Large Language Models.

Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction A Survey of Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T00:49:58.910725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:49:58.910725Z digest=sha256:7f83316a55bdf6615aa58628f26123c6703ad88703707200c8a23d1b0e9e8564

Pith citing papers

Observation 74308e21-8960-496a-9d51-81d753f7543e · inbound

Factored Causal Representation Learning for Robust Reward Modeling in RLHF cites this paper.

Factored Causal Representation Learning for Robust Reward Modeling in RLHF Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:20:13.564528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T14:18:33.768962Z digest=sha256:bd83087a1ccef519f4643ef14622e5d69194d9856d663b774e93cc6699fc971c

Observation 90bf14c4-aaac-4d72-a5a3-94d7926aa0ef · inbound

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure cites this paper.

Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure Causal Reward Adjustment: Mitigating Reward Hacking in External Reasoning via Backdoor Correction

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:24.242250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T12:22:06.635622Z digest=sha256:d519f1b66564812cf19c36d628dbfcc359c42760c0f5ddbf9ca98562b647508a