Pith. sign in

Paper Citation Record · LEDGER

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

As of 15 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 8 inbound Pith citation observations for arXiv:2506.04909.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.04909 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:38:05.468538Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:03.699530Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ea1967ed-7fb3-4bfb-a259-2ff5ac84da73 · outbound

This paper cites write newline.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.027941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.027941Z digest=sha256:19052ce69328b1b7bada702ea0a39a9b20a4e684817918d040b17dde08b203f2

Observation e7f2c0be-31e9-46d6-b0b0-ecc9b43caf03 · outbound

This paper cites and Mitchell, T.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Mitchell, T

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.795250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.079573Z digest=sha256:b80b61c490880992f8bc2e7b21b56c822842411dc9d35e4c7d371614ee141ad7

Observation bc0d4022-cb2d-464d-b884-7197d6d32dcc · outbound

This paper cites Discovering latent knowledge in language models without supervision, March 2024.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Discovering latent knowledge in language models without supervision, March 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.698680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.247102Z digest=sha256:cdc170f8805fb157d72edbcfba850ec0181b5e2bb2251d59c2502d998c864a49

Observation f8cea9d1-5ff4-40ad-ad5a-f045ea23bb2c · outbound

This paper cites Localizing lying in llama: Understanding instructed dishonesty on true-false questions through prompting, probing, and patching, November 2023.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Localizing lying in llama: Understanding instructed dishonesty on true-false questions through prompting, probing, and patching, November 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.506436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.364755Z digest=sha256:76090bf32b8ccfd5f97b1202f6c5b860cbb1adee786efab7c8dbceb20b5cb7ad

Observation 9bd5f315-7111-4238-8abe-a69eb08545b4 · outbound

This paper cites A mathematical framework for transformer circuits.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models A mathematical framework for transformer circuits

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.444405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.444405Z digest=sha256:aee43ee44de9d27a6ad7b614bce27f8be707f9e9751aa6ae4af3eaf705389776

Observation 2276cea3-94aa-4d69-a342-392a370dd943 · outbound

This paper cites R., and Hubinger, E.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models R., and Hubinger, E

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.264957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.535294Z digest=sha256:96b59e8bde3877a8a3f64857cf18f7a3ca16dfe0d84cfb7157bacecec85e099f

Observation 7a6dbadb-b282-4992-87e3-ce9132f2a1b7 · outbound

This paper cites Deception abilities emerged in large language models.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Deception abilities emerged in large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.688962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.688962Z digest=sha256:6736f58a23cc6cae2589d56a50169a6111f6fcf5689cb2515c0c65fff201517d

Observation 2ff6f362-bac0-4344-bdde-0a39ca7bbecf · outbound

This paper cites an unresolved cited work.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:38:08.038259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.814492Z digest=sha256:fdda496795f6500b0fe7a71ad1832683ed4a3baa759c96c59fad0b65122f3d26

Observation e43df03a-f2fd-4914-b20e-501298c205bc · outbound

This paper cites Y., Song, S., Hajishirzi, H., Kornblith, S., Farhadi, A., and Schmidt, L.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Y., Song, S., Hajishirzi, H., Kornblith, S., Farhadi, A., and Schmidt, L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.930957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.936643Z digest=sha256:b880a7086cfbcba7a787716da27797de8603138c46aaf41707db17734c6b2ff2

Observation eb049e20-912a-435a-9dc0-7d6644527c2b · outbound

This paper cites Large language models ( LLMs ): Survey , technical frameworks, and future challenges.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models ( LLMs ): Survey , technical frameworks, and future challenges

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.057101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.057101Z digest=sha256:45bcd2ba8b670269135801f9ee126947c12fd16e49d56cde0aa28f108bf62b5b

Observation c7fae250-c8c1-4559-96df-6ab28f368dd5 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.263205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.263205Z digest=sha256:e675d96549716261d682fd011c29e27c23b1627abc1bf860601a28fb9fc61468

Observation eaec6f86-1555-4ac8-9a0c-172a61f328ca · outbound

This paper cites Cognitive dissonance: Why do language model outputs disagree with internal representations of truthfulness? In Bouamor, H., Pino, J., and Bali, K.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Cognitive dissonance: Why do language model outputs disagree with internal representations of truthfulness? In Bouamor, H., Pino, J., and Bali, K

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.360826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.360826Z digest=sha256:fca0676a880efa749dfc1456781da03064f33da6b99edd19471deecee251299f

Observation 76fa1426-5849-493d-95a1-7de680df65f6 · outbound

This paper cites Faithful chain-of-thought reasoning.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Faithful chain-of-thought reasoning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.798430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.478408Z digest=sha256:e0b1c59b2e5a5af5a02f8f39883699c2b971d4fa0589671191d13391f2a37600

Observation 04ccd2a7-7794-45a0-8e7f-e4ce61697f42 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Frontier Models are Capable of In-context Scheming

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.666723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.666723Z digest=sha256:cfdec35010d7f8cabeb8c8b909e13dda48b2465c988a37a559463ed9074ad7f3

Observation e39c8e0a-756d-447f-9f73-173382954953 · outbound

This paper cites Locating and editing factual associations in gpt.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Locating and editing factual associations in gpt

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.670229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.799399Z digest=sha256:49610c596b4b2274a14bc380bededb2f46289a4bfea4e4f57a8d331777ea68b6

Observation d1ed4703-bad2-484d-8e89-88bf35bfc607 · outbound

This paper cites Interpreting gpt: The logit lens.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Interpreting gpt: The logit lens

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.579925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.887928Z digest=sha256:f5731cbd22e757a54e58b71f3646ea2592b60e8d0c4b4210eee63e3a04464bb7

Observation 31df4916-8878-494f-8a3b-af405b58173a · outbound

This paper cites S., Goldstein, S., O'Gara, A., Chen, M., and Hendrycks, D.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models S., Goldstein, S., O'Gara, A., Chen, M., and Hendrycks, D

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.443129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.007213Z digest=sha256:23fa2c6fa01a522cce8ed309a0ddae59a16eeda3e06ee8221a8701116cd19884

Observation b9f1ec61-4e4e-4481-ae94-4bd3ba97bcaa · outbound

This paper cites Large language models can strategically deceive their users when put under pressure, July 2024.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models can strategically deceive their users when put under pressure, July 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.324772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.146066Z digest=sha256:22c76a7a0035bd9f4f955c2d289c1cfca66a36542567f4426c1640de602c042b

Observation 218aef99-cf33-4e90-9326-c7eadb5a5ed0 · outbound

This paper cites an unresolved cited work.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:38:07.182106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.241493Z digest=sha256:61034b04bdd09384c8afd86d5a84a83f2c8d8492e3d7e90f8c058b08c1b2c5b1

Observation 72b9e771-10e1-463a-913d-1715f16c41cb · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.052683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.365502Z digest=sha256:29878e97012c51448b55a537a3c3ff23515cbcec2684920de9b8b98714d7fd15

Observation 45882e9b-bb11-42a8-8948-87676ff49ba7 · outbound

This paper cites L., Sharma, A.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models L., Sharma, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.977885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.499667Z digest=sha256:a0f9cf3a1e8f5fdfdc7f84cee17200fdfa5f91217ee50d39a377c62e3f6d9278

Observation 16202906-50d5-480e-8771-259ecba3b2d7 · outbound

This paper cites M., Thiergart, L., Leech, G., Udell, D., Vazquez, J.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models M., Thiergart, L., Leech, G., Udell, D., Vazquez, J

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.634111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.595514Z digest=sha256:aa74a2fa383fffd29af752019e5cadce05ed4c0c00fe0c0e67d2651b6f12af50

Observation 5b7709c1-7bc1-44a4-a501-4270a70d1004 · outbound

This paper cites N., Kaiser, ., and Polosukhin, I.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models N., Kaiser, ., and Polosukhin, I

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.707675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.707675Z digest=sha256:702fdcde16a7c6746f9e746673b29decc42dc872e81900d3279d5e67aec49eac

Observation a4a24f4c-66a9-4b16-83dc-98d75501138b · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.803196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.803196Z digest=sha256:b20843155df0c7492de1eaf32f41d4f1a8ab910db8a3f50d30c8bb881fa6a98a

Observation c5497388-7067-482c-a3f7-2173f61acfe6 · outbound

This paper cites V., Zhou, D., et al.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models V., Zhou, D., et al

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.941743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.941743Z digest=sha256:9833b8d5b018dab19323b8e0f3c2aeb2a6950c3a35f732bd9c484178e645f2f7

Observation 0a8dbc01-e3a2-4e56-9687-20e072740ba3 · outbound

This paper cites Qwen2.5 Technical Report.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwen2.5 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:05.057064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:05.057064Z digest=sha256:8e7c1269135399b2fb96ae30bdb048b516ebd238a5238d0fa720c917662d9631

Observation edeeb8da-e7d3-422c-b360-568a11f9766b · outbound

This paper cites and Buzsaki, G.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Buzsaki, G

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.305864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.173034Z digest=sha256:111e417d58c0938d9718f888e99b1f0dcb6e38b7d680f229ee252cdac34bc9a7

Observation 9d6147ff-ac8b-4ef5-a946-2a7379e31e15 · outbound

This paper cites Reasoning models better express their confidence, 2025.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Reasoning models better express their confidence, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:05.292046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:05.292046Z digest=sha256:296ee0a3305fdf6dfeaef7089891fabab6c96b9d9acc91c75803fdf54545854c

Observation 2acd94a5-b29b-4172-b401-4f472583fc24 · outbound

This paper cites J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Kolter, J.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Kolter, J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.029672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.352408Z digest=sha256:687bc06addc34867f2e2f7b4463f2f42af74b6ce3d996a1c3747f47da0a95724

Observation ce335cbb-4e4f-4164-ad63-f60edd888e08 · outbound

This paper cites write newline.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:05.813873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.468538Z digest=sha256:fc8a0bd5ea7097a37c3055b8f926be643b2ddaee3b3bca153ae415039d07c28f

Pith citing papers

Observation dc1e993c-71e8-4590-b811-d4843a5ebd9d · inbound

Security Concerns for Large Language Models: A Survey cites this paper.

Security Concerns for Large Language Models: A Survey When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:03.699530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:03.699530Z digest=sha256:60cfcbd3765bfffa97b082cee7fe5278d011aad209ecd8c6464c9de4e1bcc0b9

Observation d94b862f-9d0a-46a8-8e48-25a4526a7381 · inbound

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers cites this paper.

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:00:51.707342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:00:51.707342Z digest=sha256:5069bdd38ea67f010b3378238ae167841a28a1c7d9b67c30f04be0ace7cd11c6

Observation 0c106958-39d9-4aca-85ba-031dcddc8406 · inbound

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs cites this paper.

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:45.041136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:52:45.041136Z digest=sha256:5ea0530a7753d45aedd1a72d7b2f8305e97e99b5e6b2d9281160a1e7ee237e2d

Observation 5d7330bb-6195-47f7-9d55-0d3ac29e083c · inbound

DECOR: Auditing LLM Deception via Information Manipulation Theory cites this paper.

DECOR: Auditing LLM Deception via Information Manipulation Theory When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:28:05.389621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-20T06:27:10.445757Z digest=sha256:e89a378019a91b2a34cf929cb4bc6b5ae5333879808e6fc417126ec08d769b7c

Observation e4811b80-86b1-48d7-beab-a368e862aeaa · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.716480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:6a85773b32d3426c526d68afd4cbcc00faa1944e06630b016e5aa573e5b9218e

Observation 88499384-a89c-4608-a2f1-9f5addf98a37 · inbound

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates cites this paper.

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:08:07.647083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-07-03T13:02:46.485260Z digest=sha256:b04b6b1439528bd39a4e01ca7e48cc7a1c4d3acb75864f7c1e6b33777343e85e

Observation 54f538bb-6061-47d0-a0c1-f4f38945e484 · inbound

Transcoders for Investigating Deception in Language Models cites this paper.

Transcoders for Investigating Deception in Language Models When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T01:06:13.525747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:06:13.525747Z digest=sha256:c1cdbd81afb7d5f128580c30cc82ff1f4919e0ed3cc43004079964fa3786f8e1

Observation a8ada123-d371-4e1c-99e8-b656e1ef1880 · inbound

Risky Business: Measuring The Faithfulness-Safety Tension cites this paper.

Risky Business: Measuring The Faithfulness-Safety Tension When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T13:40:57.774902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:40:57.774902Z digest=sha256:595eee5d1424085628197f84794566e22de435747cf4900787bbb755ccbafbcb