Pith. sign in

Paper Citation Record · LEDGER

Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2504.05419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.05419 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:15:27.776270Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 666e2612-48b2-4d5a-83da-d9edd88670b6 · inbound

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective cites this paper.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 130

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:12.982125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:12.982125Z digest=sha256:cba6bff5e47d4132e9475ab67b5e0a33a1f1143424c8525be5830cdc8eb2b2e5

Observation 2532eceb-f695-4c11-b079-920187603302 · inbound

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency cites this paper.

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:20.228761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:20.228761Z digest=sha256:191b35a42a69ba27a9da67b07dda884c84e8e88424da1a910bc8548404c2bdcd

Observation 8a2025f4-41bc-4e8f-b31d-7ca5b723881d · inbound

The Geometries of Truth Are Orthogonal Across Tasks cites this paper.

The Geometries of Truth Are Orthogonal Across Tasks Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:12:50.596167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:12:50.596167Z digest=sha256:b49b971a6969493613ef7b2348cfef7b28b632ddd5c652a0522fdc4b62868941

Observation 2a21f87a-3f81-4075-8720-f0df015b0e62 · inbound

Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement cites this paper.

Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:49.924159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:49.924159Z digest=sha256:63d0196cbc870eb06827573f916050ac2e0ec7fdd3880086f8d65e203271568e

Observation 3936ad77-716d-4947-a689-efe6d943f63d · inbound

Real-Time Progress Prediction in Reasoning Language Models cites this paper.

Real-Time Progress Prediction in Reasoning Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:09.871354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:09.871354Z digest=sha256:063bb1015ab5026d82843839549db677d24a5240342e81becc12f5f684f6e57c

Observation 23c1314f-b0f2-4e6e-96f3-05dd4da8282f · inbound

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs cites this paper.

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:31.411480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:31.411480Z digest=sha256:29deaf06a8bacfa1ca4347afcca3da95a1c4420f0d31fdc6aae4a2fd78dcce12

Observation 1e65bdff-a7a5-4a94-bb64-b63464b4fe62 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 237

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:18.051568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:54:18.051568Z digest=sha256:9fc3f0d91e932bad51a209b4584aabc43b475991406a2e6bb2652dcc0f251127

Observation 0c49d9c9-e8d3-4485-9753-cb07cd765495 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:03.846399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:03.846399Z digest=sha256:d707e2e8ee75d6c9f9dcd834984d4a8b9b1b6e3689373bb2cf51fbb801729e40

Observation 8d6fa640-ea61-411e-a4f8-da6b9361be6a · inbound

Rethinking LLM Parametric Knowledge as Post-retrieval Confidence for Dynamic Retrieval and Reranking cites this paper.

Rethinking LLM Parametric Knowledge as Post-retrieval Confidence for Dynamic Retrieval and Reranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T23:34:30.832661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:34:30.832661Z digest=sha256:1141290053d6c9f37730d70e8549cb4cb616fb602a9e182f11e09ff35b622194

Observation 768c89b5-0e74-42c4-865e-5124e83f0e87 · inbound

Entropy After </Think> for reasoning model early exiting cites this paper.

Entropy After </Think> for reasoning model early exiting Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:52:35.578407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T11:51:58.579048Z digest=sha256:b4f0e9dac31206b2e2056d5589cfb449b23fed5fe76ce920673a34f43e78c80f

Observation c9db14a0-1101-485b-ab7c-eda414fd952a · inbound

Learning More from Less: Unlocking Internal Representations for Benchmark Compression cites this paper.

Learning More from Less: Unlocking Internal Representations for Benchmark Compression Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T06:01:59.124009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:01:59.124009Z digest=sha256:8d988ddedb4a95d0224dc6d383c55ccfd78edb7190bf8d13a7370f42282b4742

Observation 47325669-6598-4559-a7be-eaa5878c0911 · inbound

Conformal Thinking: Risk Control for Reasoning on a Compute Budget cites this paper.

Conformal Thinking: Risk Control for Reasoning on a Compute Budget Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:47:32.666284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-16T07:45:53.184670Z digest=sha256:662dc360b6e5ba9b60585b7964edc1c0ae5d58d4f7b2ae8059fb2f23d2e7cd42

Observation ded4de4c-6ae2-4e38-aba3-a867ef938e09 · inbound

Emergent Manifold Separability during Reasoning in Large Language Models cites this paper.

Emergent Manifold Separability during Reasoning in Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:35.060139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T20:12:05.809065Z digest=sha256:3a486cbe1f23be46795b6b76e56778a123d6049edd28698c443e3e1deabdc2f2

Observation 0cc44ec1-660b-4072-b918-96deb8804ab0 · inbound

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality cites this paper.

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:35:51.917676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T18:26:04.301213Z digest=sha256:9745d09c2ae718118fa6e8356f77962fe5cf4f933b20c139b38a83bb98929762

Observation 2f2e76be-f53a-4869-be90-2dc30e3354a6 · inbound

LLM Reasoning Is Latent, Not the Chain of Thought cites this paper.

LLM Reasoning Is Latent, Not the Chain of Thought Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:53:04.614744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T08:49:05.178087Z digest=sha256:ef9e9b37ae7fa61e621de1aa708ec0067fb8b46d76808b608d70996bd1f8c557

Observation 5f202c30-32b6-46dd-9e5f-9b5578072dcc · inbound

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping cites this paper.

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:29:47.262377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T00:26:45.372232Z digest=sha256:869427aee1c87ea1b737bd4a86ead0d5c9022e59ffe1864b1b34808b0208a49f

Observation 12528287-ffee-497e-90e6-4cd667937eb6 · inbound

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation cites this paper.

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:10.371417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-08T06:38:32.413172Z digest=sha256:516593a12c07b47c511acd54832104da035d36d53bab2e47f2ef55e8a687f238

Observation 72c7a5b4-7de5-4919-a122-727a24f18f64 · inbound

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation cites this paper.

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:05:40.207092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-07-01T09:56:48.019404Z digest=sha256:2e9332751a94521690b6b07fe084c55858d902c683ef83e82291bbf83f95e5b5

Observation 30b3732b-6007-4ad6-8a90-e05a311398f9 · inbound

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models cites this paper.

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:21:08.888017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-09T17:20:19.586214Z digest=sha256:40f7ffd6fa2fa3c602f5c2fa5ea2f7ef15f4076d42195a2b3bace6dc64d604f7

Observation 743a19ff-25c2-4fe5-baf3-d38cf596103e · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:56.545726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T00:54:25.549158Z digest=sha256:68a2ac31474b1be52ecf05554d8e207c859bf85add44ee82a8f020f2f5684adb

Observation ba999c67-75b8-4b6f-8675-af4b50895d01 · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:29:52.735984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T08:29:09.122055Z digest=sha256:4d56a3da8c002267f17341937d0c1fb2be6856495eac7f563a7ea58281ff0e2f

Observation b83c8184-7213-4655-b137-7250c9e2aee1 · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:10:58.214344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T01:54:34.000216Z digest=sha256:53eacb3f14bebe3520640bb1ec7f22a1b6725725ef6f532d1f3bec628a54544f

Observation 806489af-17ad-4a9b-9c82-2aac09eb005b · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:12:28.383220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T07:09:02.672233Z digest=sha256:6ee4af1c86143d6dd3198348981ecde1500feefba8c1491cfadcbf207a88b317

Observation f2733e87-406e-44a2-918b-26f98d3f87b2 · inbound

Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal cites this paper.

Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:31:25.390664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-12T05:11:51.012343Z digest=sha256:449835cdac816ba82d766d55d8017fe102fa758fce96fd5f3f5b87b8f51839f7

Observation b626de8d-581c-4f7c-9b28-4eaf8ae8f688 · inbound

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking cites this paper.

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:46.315000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T03:58:46.417344Z digest=sha256:3fbd0ed8997e0b9e5c267a0891d8b68f1de44cdf92dcba1ef38b412633e95d34

Observation 859c5d38-85c4-446f-92dc-24b86657a5ae · inbound

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking cites this paper.

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:12:41.168062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T17:10:13.032423Z digest=sha256:37884c9e26da638d0ce41c38be15fc7dc2ea480dd948f88ed1e3b6fd058f91a7

Observation 7e90ac0e-dfd1-4d35-b8dd-28d61f644204 · inbound

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards cites this paper.

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.363599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-29T22:58:46.313028Z digest=sha256:c96a4fef3ef714000ee6062fd01a10861a54b2622044938f5910787e6a4d2875

Observation 6db0d441-f816-455a-b782-27e796065c49 · inbound

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking cites this paper.

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:40.955133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-29T16:59:25.127619Z digest=sha256:0b6d3c38db503b4d30a031ae5213500d470cf95c520bfa09f9fe9f714ace2836

Observation aaf993bc-7338-4ca7-a8f8-b87c3cca3282 · inbound

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling cites this paper.

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.894680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T10:25:10.559953Z digest=sha256:6eb48ee8f6309035cdd440c2c63a91b26ba9dc2a5e37dfe0d1e9a5e2e8c71895

Observation ae8514c4-2188-4734-91ae-0a0933d89460 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.591198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:105f53374f679011410741b0650e83d46188fcf448103300c004bcab7a6a3303

Observation 9c75a31b-a1b5-4f76-93b5-bce026bddc92 · inbound

Dynamic Rollout Editing for Reducing Overthinking in RL-Trained Reasoning Models cites this paper.

Dynamic Rollout Editing for Reducing Overthinking in RL-Trained Reasoning Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-27T01:20:20.414019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T01:19:55.164835Z digest=sha256:9b08cb5b151c632cf9c7c5ab4ca18ee750a26d292af0765ea6f29e86df0f1a7f

Observation 1e05eb4e-4a2a-413b-874c-35d26c5d29d5 · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 111

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:38:55.839168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:dd6d6107a00a365e2d000dabf8299a6d3e6177b457d36afc280a3ac12266f232

Observation 3c41cc19-79b3-4f5b-baec-6e243190baac · inbound

Does the Same Token Mean the Same State? MoE Routing as Signal for Reasoning Control cites this paper.

Does the Same Token Mean the Same State? MoE Routing as Signal for Reasoning Control Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
malformed identifier
arxiv_id, observed 2026-07-04T10:19:47.420429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T08:59:28.038685Z digest=sha256:643dae60535506fc7fed9e4492d56a0e7cbe14cd10198d887c024bfbab10ea01

Observation 4dfb2210-428c-4b74-aa3c-6a15235cbd3a · inbound

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents cites this paper.

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:49:45.696278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T08:32:24.949332Z digest=sha256:61e5539cd3fe34b7b84ef1cf2f26e55b10192f2036998a611ae3c876169694c4

Observation b90db3b7-ca9b-40f9-8560-2806f98cedc5 · inbound

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents cites this paper.

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:49:46.404977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-26T08:27:28.236073Z digest=sha256:aa60ae3a69e539ab1933d9d724688e2f6523dbdd1e9ed7161597e133e4112e27

Observation 6305eb4e-aec8-4061-8d2d-f359e683fe93 · inbound

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning cites this paper.

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T14:00:54.863716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:00:54.863716Z digest=sha256:0d98db1b8d5a0fc92dcb0c6c0a07f5288f18bdfb62c9c2100e4f9568c14bf368

Observation 499c5b30-9477-40a3-a998-f70e2869604d · inbound

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning cites this paper.

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T07:28:23.769012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:28:23.769012Z digest=sha256:8672c41dfa1afbfff88c063951a02593a30c935c7ad42e326ec73a4c2f623d1c

Observation 8ca5437a-6d42-489d-be36-7413dc947a89 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:07.720174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:07.720174Z digest=sha256:23ab6a284fb0350bd129edf1ee57ffba6d720b031fdf0e0493c809c1f09a2f0d

Observation acefe7b9-9aad-474f-8d33-c81a64081175 · inbound

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes cites this paper.

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:49:36.297426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:49:36.297426Z digest=sha256:2d48719474639c4ba6be2ca5109aadb4f73868ccfdc71ff65d1cd21a5b2f6c48

Observation 344778ce-3ac8-441f-9c03-a599cab63584 · inbound

On the Robustness of LLMs' Internal Representation of Code Correctness cites this paper.

On the Robustness of LLMs' Internal Representation of Code Correctness Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T00:15:27.776270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:15:27.776270Z digest=sha256:3370c60cc6dc04dfbe99242c7de7b655f771786498faf6f4b0cdc18862a44f43