Pith. sign in

Paper Citation Record · LEDGER

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings

As of 19 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2608.04735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04735 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:51:52.279090Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact6
  • verified fuzzy17
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7847c1fd-240f-4c1b-8b72-b44e8ba17873 · outbound

This paper cites Inspect AI : Framework for large language model evaluations, May 2024.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Inspect AI : Framework for large language model evaluations, May 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:57.025625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.828961Z digest=sha256:df8340d5300aad33bcab48b7cd59c07ef5f649eb10251f594c4b32ac4f164556

Observation d2be7f1d-5e18-467f-9912-e82813a20e15 · outbound

This paper cites Introducing Claude haiku 4.5, October 2025 a.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude haiku 4.5, October 2025 a

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.835810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.876010Z digest=sha256:3f4e56b28433951a0f14769a8d31f564513787708f78ec7728ed5fecb4cbb016

Observation e95b416b-de1e-4fd8-92ee-637b4e2811a0 · outbound

This paper cites Introducing Claude opus 4.5, November 2025 b.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude opus 4.5, November 2025 b

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.491211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.923038Z digest=sha256:328a754642cb2403938f3ed4eb266504dfe3bead6323e693318ec8ece18f6cf5

Observation 1a90898a-bc18-40f9-8b9a-f0032c091536 · outbound

This paper cites Introducing Claude sonnet 4.5, September 2025 c.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude sonnet 4.5, September 2025 c

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.238274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.968522Z digest=sha256:a381177d0d873d3f77f0749bde8240a9a6c1fd04f60f8c1867bd1d89d129567c

Observation 96f401f5-f452-466e-8c89-fc3e111287e3 · outbound

This paper cites Chain-of-Thought Reasoning In The Wild Is Not Always Faithful.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.013543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.013543Z digest=sha256:f3dfc75ebf00a339dbd0f0cbce05d809e739cf086b096ebe69a1ce6e193c65c0

Observation 366a74fc-0792-4b39-b5c3-c60280eb4461 · outbound

This paper cites Biases in the Blind Spot: Detecting What LLMs Fail to Mention.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Biases in the Blind Spot: Detecting What LLMs Fail to Mention

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.079079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.079079Z digest=sha256:83401078afae3ee7baf273108022e808cf94b732c33491552773b590475a53c3

Observation 3b791013-7f29-4d0b-888a-19e074ff7dea · outbound

This paper cites How does information access affect LLM monitors' ability to detect sabotage? arXiv preprint arXiv:2601.21112, January 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings How does information access affect LLM monitors' ability to detect sabotage? arXiv preprint arXiv:2601.21112, January 2026

Reference 7

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.773964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.133048Z digest=sha256:2c89f3eeb109a1bf43477f07d0b3348c9d4f2ce20b7d46f2d0d4035ec2b17cc8

Observation e411a152-ded3-44c0-bd5f-77beed122afc · outbound

This paper cites CoT red-handed: Stress testing chain-of-thought monitoring.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings CoT red-handed: Stress testing chain-of-thought monitoring

Reference 8

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.696529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.189440Z digest=sha256:c0fc3ab3c0daeabf269ac686bdd98e12861beb876a669a9ee61c66bffdb4b1bd

Observation 0aa7a58c-ad9e-47a7-99cb-42408b2e00b0 · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.240956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.240956Z digest=sha256:1686af7861d046886499d2eb220b3d1695fea25ae40e727d8dcb4948a5c4f7d3

Observation 791437e5-9cf6-41f3-beb7-14463f28461b · outbound

This paper cites Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:51:53.503854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.301008Z digest=sha256:38ee1149c1268040bd08b9a49d6aa49b3f2b66cd7de51aec121e84acc49fac86

Observation ee052de4-fd50-4743-b7dc-1918bc7d93d5 · outbound

This paper cites Censored LLMs as a natural testbed for secret knowledge elicitation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Censored LLMs as a natural testbed for secret knowledge elicitation

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.622765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.364168Z digest=sha256:5b8b09b35d11f00ade26c20cf177b6cf75bb3a1866f002789ec90e55560d536a

Observation a1af007d-4c82-4494-8810-dfe623e17938 · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Reasoning Models Don't Always Say What They Think

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.482929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.482929Z digest=sha256:9bfc0bb8e1c3d103a5a13d2b25f622204451983c7b00ed95930b34bdb69cdfae

Observation 31bd8a52-def3-4980-b486-f4912b8a29b7 · outbound

This paper cites Lee, He He, Ian Kivlichan, Bowen Baker, Micah Carroll, and Tomek Korbak.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Lee, He He, Ian Kivlichan, Bowen Baker, Micah Carroll, and Tomek Korbak

Reference 14

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.542529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.556485Z digest=sha256:f72de47f8a7c1403a2494a269dfe77941c149b034ce805a16cccfb9546c48242

Observation e5e97ee0-eae1-4eb5-b90e-08d1df8d9865 · outbound

This paper cites Are DeepSeek R1 And Other Reasoning Models More Faithful?.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Are DeepSeek R1 And Other Reasoning Models More Faithful?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.627288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.627288Z digest=sha256:8c5eb591991bfe546672655852f3ab29f1714a24dc8ee8b3b147849c4f1ff971

Observation f3d8dc31-19d1-4cda-9ead-c45b620d98b6 · outbound

This paper cites When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.719517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.719517Z digest=sha256:c54f03eeb0e028b22956735959ebd0ec788a9c360052781616c4c442fdde2f27

Observation e76d3e12-a314-4abe-94a4-1407cb4be026 · outbound

This paper cites (some) natural emergent misalignment from reward hacking in non-production rl, March 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings (some) natural emergent misalignment from reward hacking in non-production rl, March 2026

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.031994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.834128Z digest=sha256:d2b885c6be8de706a98215e809735fd8cb6b7b9d2f820b69a2bf0573be0fb25f

Observation d3aea62b-7a06-4803-9dad-47dc30961199 · outbound

This paper cites Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y

Reference 18

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.467492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.946000Z digest=sha256:30ee77714d92bfe1455a6966aac076c1c1893c683dc48279588073bf9663b259

Observation 8cb673e0-48bc-460e-b3d1-bed4380fa23f · outbound

This paper cites Noticing the watcher: LLM agents can infer CoT monitoring from blocking feedback.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Noticing the watcher: LLM agents can infer CoT monitoring from blocking feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.058436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.058436Z digest=sha256:9a53155310f9f8a165508cca76c38608ddeff166828c4395c8d35458edeba9ae

Observation 3e760a5c-06ea-4874-9859-3c1560cfed5b · outbound

This paper cites Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.127591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.127591Z digest=sha256:76977c9103f6768d11ff655afb5f437c7458804d714eed050432db60566182a7

Observation 22c44c1e-22f5-4d04-b6fd-fce58ce9fa4f · outbound

This paper cites SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.183344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.183344Z digest=sha256:9906e5db1c438a255a9ce44c51028eddd63041eb0fa30b876b794d7fe3360fd0

Observation 1876190b-3bb6-4f5d-8e47-bda706cf4f6a · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.186621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.186621Z digest=sha256:5e7c1193429dba3181d129948087304b757ae48015711dcd10586eb9fac54fa5

Observation 1c2634c9-8748-435f-89fc-1556adf7a8e6 · outbound

This paper cites an unresolved cited work.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.188944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.188944Z digest=sha256:44b4aa03d66ffb4d68f74e04b884bdf8956331dc454f5de3349108d47fc37cd5

Observation 3651c42c-9b21-436c-ab27-1bab369e683f · outbound

This paper cites Kimi k2 thinking, January 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Kimi k2 thinking, January 2026

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.762764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.191108Z digest=sha256:3d1b73b16a6071e41772719ac694e3da4e6ec41bce2e787414b73d2def50be5c

Observation 4760181c-3f11-4e6a-8253-4d579d9e9421 · outbound

This paper cites gpt-oss-120b and gpt-oss-20b model card, August 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings gpt-oss-120b and gpt-oss-20b model card, August 2025

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.503751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.193355Z digest=sha256:2bc768c60f0ea41f215412488b7cac2484bebeb14f7752664e251c379479ba3f

Observation d17cf599-3150-4a5b-8cc7-14591bd740eb · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Llama 2 via Contrastive Activation Addition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.195387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.195387Z digest=sha256:de3ea7fb32ae6c0f8c01792800f907078e8bd6760985c0fc422bd0c721df2ee5

Observation 4269579b-a6a1-498b-8a71-ec6210044950 · outbound

This paper cites Large language models can learn and generalize steganographic chain-of-thought under process supervision, 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Large language models can learn and generalize steganographic chain-of-thought under process supervision, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.200090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.200090Z digest=sha256:02dc43b66561f6fed950d9588f1399c1659f3b77335d0f46fe6fa31308e5ae6c

Observation d19fe64f-252a-44d0-be19-488d429369ef · outbound

This paper cites Steering Language Models With Activation Engineering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Language Models With Activation Engineering

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.202208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.202208Z digest=sha256:62627181c7c3ebb1ebb028cb526a867ec4ec79d47585ae4312c01f0df690905c

Observation 1e354193-0b20-4880-bdd3-940b439a681b · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.204689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.204689Z digest=sha256:84fde04858b8ff2f51e20df856740de928f7bb71826a8be0f0d495ed6578983c

Observation fc5cb71a-18b1-49eb-a2ec-5d8a375928a1 · outbound

This paper cites Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.207562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.207562Z digest=sha256:13e1f7cb96b245967e888271edc358f0dc77b03db3c9ac7072096ff6ca95bbce

Observation d43bd6ce-84b7-4f2f-b8b1-aa28a045cfa8 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.209806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.209806Z digest=sha256:4dd71e57649632dc2d4729e5ae69181a98b985eb827bae633dd025e3af7fedf1

Observation e8b455d8-8d05-46f3-859c-f38727e2e722 · outbound

This paper cites Grok 3 beta --- the age of reasoning agents, February 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Grok 3 beta --- the age of reasoning agents, February 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.227558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.212330Z digest=sha256:60e293af2cf9d9d95cfd0f8e3af6eaef5bc98672d1c832143c5a13f4bc8a6b5f

Observation 54fed9a0-8407-4395-ae8a-33038bcb0cfd · outbound

This paper cites GLM -4.7: Advancing the coding capability, December 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings GLM -4.7: Advancing the coding capability, December 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.965790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.214472Z digest=sha256:c9628435d55b56fe28abe167702e2b7475022170dac6ece372c87d1ab43108f5

Observation 13fb16fc-05d5-414c-8ed0-4e4729999ed4 · outbound

This paper cites Can reasoning models obfuscate reasoning? stress-testing chain-of-thought monitorability.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Can reasoning models obfuscate reasoning? stress-testing chain-of-thought monitorability

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.216459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.216459Z digest=sha256:eeec532c6e964f47284afd1e00b487399b34cbfc0f3398b79a602a5da0166c89

Observation 2437502c-1a34-4113-ae1e-bba68ec7ea31 · outbound

This paper cites 2023 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2023 , eprint=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.218672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.218672Z digest=sha256:289ac41a358df58e0814ea8c7e0b615a58544818faa6ac1664e2b1c109c58d0c

Observation 927958a4-176c-4630-b2be-86610271f78c · outbound

This paper cites 2023 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2023 , eprint=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.220659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.220659Z digest=sha256:b1aeb548848ea60b16b2e1db3063dfb1d30c89de0c3ad76c0d228afd5e78624e

Observation 98c666c8-bd05-4431-8064-c1bb5dc17370 · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.222703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.222703Z digest=sha256:edae68e926df5d25ebdaf59e466989d25f53ecefd8e5597c4fb7e9f14aa765c6

Observation 4f9459d3-121b-4223-a87d-8015784b7ea4 · outbound

This paper cites 2026 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , eprint=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.777138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.224969Z digest=sha256:88f4b1fda4e741c0253bccc9ad4c2af7635959267025241addbd6957491957db

Observation 5178187b-cf63-4f2c-8935-92370c368a2d · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.665282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.226927Z digest=sha256:b32ce5f5199a0ef64b4a80cd6b2707868009186ab4535adc12b0c761cd4131ca

Observation c9b6f91c-e608-4553-900d-9d7171c2293c · outbound

This paper cites 2026 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.582474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.228836Z digest=sha256:b1449dd19e1226356b9ba0e78e363001393b8f5d196992ef17f45ebeee39cdca

Observation 8a77ec12-81e6-421d-95ab-c5a11106bdca · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.230732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.230732Z digest=sha256:c5fe5f34200dc82d33d5b829dfc4bcdc6cc65cc6f1e38b3822be2c8ec3ad464c

Observation 2216db55-6d18-447f-abcf-399d79cf8587 · outbound

This paper cites arXiv preprint arXiv:2512.18311 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2512.18311 , year =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.232686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.232686Z digest=sha256:323ec38c15932a9607e0697d8721b627667384edffeff9a25096fdd491b72c32

Observation f2441273-de61-4d74-98c2-3bd6b0405e43 · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.234780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.234780Z digest=sha256:253a9694acbc78e0a2f98833f4bbd6f02d132a89cfcc68214c51400d0e660a5b

Observation 03053e1a-46ae-4826-8f62-66d4aa37aeb3 · outbound

This paper cites arXiv preprint arXiv:2603.05706 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2603.05706 , year =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.236856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.236856Z digest=sha256:36429c639d80e4186a151ccc119845221f69bcf6329fa7f0a247f7a43799fb76

Observation 73fad7a8-57a9-467f-9bc2-e574c9639c7a · outbound

This paper cites arXiv preprint arXiv:2510.19851 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2510.19851 , year =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.238875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.238875Z digest=sha256:48019b9457eebfb9f9d71a11f9feff674681886740e589fa96c1659a5bad9e0e

Observation d596953b-7e25-450b-b47c-9a8a9b4bc214 · outbound

This paper cites arXiv preprint arXiv:2505.23575 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2505.23575 , year =

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.241151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.241151Z digest=sha256:b09d62aa6ab67243e1cad72ec0ccd15ca4cbc51c91bae7907852d428d0e55ed2

Observation b9bf8812-08ae-4748-befb-961011b835f5 · outbound

This paper cites When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.243168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.243168Z digest=sha256:675c14a87df1de1a82fac1dd8e1e5e1cc11d356fd3162d495f753a186b591bb6

Observation b8b8928b-266a-431f-8ecc-d9792f291cf1 · outbound

This paper cites Noticing the Watcher:.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Noticing the Watcher:

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.381037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.245734Z digest=sha256:e6ec026aa249251f7e690a3db308a5cca32f3ff93cea23d821283a7627d6f0e2

Observation f77f12fa-2906-4938-afa5-a16bcb1a9463 · outbound

This paper cites SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.248062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.248062Z digest=sha256:1c0ca32b05afdc5ee05c8386b8f779305a31286e4377cec4459e5571bac7f351

Observation 9aedb683-1bd7-40c5-815c-4ffddab805dd · outbound

This paper cites How does information access affect.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings How does information access affect

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.250176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.250176Z digest=sha256:203f2eb8bc3b331a9d90ffe8684bd1aa779bb34b2c84d93ed4af5fdc82323054

Observation 9f4c68f4-055c-4a3b-a21e-2af33c9e81f9 · outbound

This paper cites Censored.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Censored

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.251993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.251993Z digest=sha256:a94cc646d70e667c4f4b4684849bee7f7927015b45aee10637132d09dd895d19

Observation 8866e0b7-52c9-4da1-818d-d74a206bfd76 · outbound

This paper cites Steering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.249447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.253855Z digest=sha256:89737e4783c87b29620c21098faa0a139a8093df96e96b16580b441b558577ca

Observation 1bde4f1c-47e6-4640-b119-782c5d80421e · outbound

This paper cites Steering Language Models With Activation Engineering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Language Models With Activation Engineering

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.255930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.255930Z digest=sha256:bff987baecf0349d6f92cc29028f2745b240c070cb063b1a4730e5b7f16a6c86

Observation 47e18b2a-f703-4c80-9d3f-be93e2d71ed6 · outbound

This paper cites Are DeepSeek R1 And Other Reasoning Models More Faithful?.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Are DeepSeek R1 And Other Reasoning Models More Faithful?

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.258144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.258144Z digest=sha256:e784739b8942408c60f262e6b20a925cc1bd003bfa8fad01e08a6bc083649eaa

Observation 27e34088-3de7-4aec-a62c-7553b5bb1345 · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Reasoning Models Don't Always Say What They Think

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.260226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.260226Z digest=sha256:dd178d1fe78b14b293a917bd65501a868bf4afe7e6a4283c4e315b89420d878d

Observation 8daac443-4250-474f-8592-87fb3f5eef50 · outbound

This paper cites Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.262506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.262506Z digest=sha256:61e67445f6460ec5d0142efd53dec82d1a87dcbb7a3cfed35b608142257d6b7b

Observation 10dc86da-4f31-49dd-89d6-1d84367a99da · outbound

This paper cites 2026 , month=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , month=

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.034479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.264644Z digest=sha256:20d4be0d808894d6981c35a5148df8383c9493e8b10cf139778b89a7162627c3

Observation 0370dbc7-3f1a-4713-a0a5-382c4d7f45b6 · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:53.801715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.266564Z digest=sha256:e4f498127e45556453489c325d9061b90d47698fed38b04331eaf559a2babbaf

Observation 10862a44-87c1-49bd-a152-b497a20cabab · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:53.612616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.268522Z digest=sha256:fdf62fe95ab7e3f9680633f02f2b00c4613f0fbb2e3c1d19c71e0e63ccb547ae

Observation ff7822dc-7acd-4f27-a4e3-83cd1b74624f · outbound

This paper cites Humanity's Last Exam.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Humanity's Last Exam

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.270575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.270575Z digest=sha256:a4c548da9d8d91c23614a3915d4a44e5e14d104bb77110f989cae5d5bc7ef71d

Observation f68615f0-7475-41bb-8659-4c087ba4a378 · outbound

This paper cites 2025 , month =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , month =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.272644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.272644Z digest=sha256:8ef3468f511bac4880782fda4faab3ca91a75bfc0a1094a6330f12ab90209a18

Observation 8e3d9c8b-8daf-4730-91e2-6918e9a77e7c · outbound

This paper cites 2024 , month =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2024 , month =

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.274605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.274605Z digest=sha256:74aabda52d6df74a1fcf7f43671e2e50fb05f7a58abf713180911982cb8f1612

Observation 5a939a46-89a1-43c6-a786-78eb546c2959 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.276967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.276967Z digest=sha256:6722f421c0216f1451ebc128fd7176ed7d6625077da14c376bba1ad7225f32d0

Observation 7975a77d-4bac-4655-b354-d47530af65b9 · outbound

This paper cites 2024 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2024 , eprint=

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.279090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.279090Z digest=sha256:e62de0e2cc9ebe700a17ae3bcbbb9d93dbfb868fd8801cec66948884a68bab02

Pith citing papers

No inbound Pith citation observations are available.