Pith. sign in

Paper Citation Record · LEDGER

Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2502.12964.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.12964 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T05:29:59.340989Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T03:06:29.134779Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7fa80d3f-1441-4d8c-a625-80d56c7839ee · inbound

The Future is Agentic: Definitions, Perspectives, and Open Challenges of Multi-Agent Recommender Systems cites this paper.

The Future is Agentic: Definitions, Perspectives, and Open Challenges of Multi-Agent Recommender Systems Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:21.890291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:21.890291Z digest=sha256:a92fc12ee4eba7cfb6fca9614fe0b072e6ece2b31d3a8dc4dbc7ad690e5778e9

Observation 29f924ff-d768-4776-ab16-9827381fb2ac · inbound

A Single Direction of Truth: An Observer Model's Linear Residual Probe Exposes and Steers Contextual Hallucinations cites this paper.

A Single Direction of Truth: An Observer Model's Linear Residual Probe Exposes and Steers Contextual Hallucinations Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T11:01:15.033088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:01:15.033088Z digest=sha256:3a97b91e4e576cf654c6f29c9b5aef0bdfccd8d485425d4c5d935a2866012410

Observation 947c7c23-9bee-4ee1-a2fe-9b64140a0c9b · inbound

Tractable Asymmetric Verification for Large Language Models via Deterministic Replicability cites this paper.

Tractable Asymmetric Verification for Large Language Models via Deterministic Replicability Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:10:42.808795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:10:42.808795Z digest=sha256:4ae4be2b7b540927fbb92457f4661afaea9e07bd4950b69c937b961a005e8817

Observation fc089923-f062-4713-9b86-731138440fd7 · inbound

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting cites this paper.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.437743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.437743Z digest=sha256:4336a8a2f18304ec4e60f0830ceaf4f4743ed26ee3610f9dfde3e9efe4d84d75

Observation fc80d791-d6d7-4c84-b149-2638c565961d · inbound

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation cites this paper.

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:53:14.123949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T22:52:58.524934Z digest=sha256:4036756cbdf1a98645a2b2bc025f65ec476f199cc78ab96c629bf8b382fd5236

Observation b3f1125c-65a0-49e9-8658-76ef33eaf6e3 · inbound

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence cites this paper.

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:53:11.900607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T19:48:13.133733Z digest=sha256:cc87cf116dfaa84b834276bc9440034470107e763e33006f16c7340addd4ea0b

Observation 88621839-a4cc-4316-bb1f-e87c28b67e2d · inbound

Complementing Self-Consistency with Cross-Model Disagreement for Uncertainty Quantification cites this paper.

Complementing Self-Consistency with Cross-Model Disagreement for Uncertainty Quantification Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.595642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:29:24.974157Z digest=sha256:7265ed99834278ed12424ee68feecb5b17f6d38e77629c707f34fe95d5d67d11

Observation 95aa4e56-a116-4659-82b2-cb422143cab2 · inbound

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? cites this paper.

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:23:24.409363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:18:36.854164Z digest=sha256:7c45d3f9355e6559cb69aa421e0ce98914397070c9e8d49f346694809ffcda7d

Observation 7b4e0c60-d32f-44c6-bb31-7e93d808b41f · inbound

Quantifying Faithful Confidence Expression in Large Reasoning Models cites this paper.

Quantifying Faithful Confidence Expression in Large Reasoning Models Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:06:29.137832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:24:38.335417Z digest=sha256:98fe1865555060d88c3641ad2dd71dab8bd43b348424a0434d24f6a4c1ef24ed

Observation da2a5e8d-e1b8-4e57-aad9-11080228da7d · inbound

Confidently Wrong: Detecting Hallucinations in Financial Question Answering from LLM Internal States cites this paper.

Confidently Wrong: Detecting Hallucinations in Financial Question Answering from LLM Internal States Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-14T05:45:08.896651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:45:08.896651Z digest=sha256:a0f0fef81e786c5c7b4b6e09eb613782a287091922b48013d20574dba5a12415

Observation 93b5537b-a67a-4993-a461-80ab3964f3ee · inbound

Reliability Scales Inversely: Hallucinations Snowball Faster in Bigger Language Models cites this paper.

Reliability Scales Inversely: Hallucinations Snowball Faster in Bigger Language Models Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T09:19:46.058147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:19:46.058147Z digest=sha256:e25cb424e4c6675ae7f63041ee49e835d552f760a4a2910e69fbd535b0f6bea0

Observation aaa44321-b620-45fd-8ab2-737f3cce57bc · inbound

$\Sigma$-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems cites this paper.

$\Sigma$-Mem: An Online Reliability Memory for LLM-based Multi-Agent Systems Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-31T22:14:36.622744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T22:14:36.622744Z digest=sha256:05687b647aa1034e269a816ab78b8b8d07acdadf6c2d779eece2558668ecae4f

Observation 09c717b8-cccd-40b1-a0a7-48359a86b0c2 · inbound

TruthLens: Object Hallucination Detection via Self-Evaluating Truthfulness Scores in LVLMs cites this paper.

TruthLens: Object Hallucination Detection via Self-Evaluating Truthfulness Scores in LVLMs Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T05:29:59.340989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:29:59.340989Z digest=sha256:1dd0036d742ee122c354686904a85f811aaeb579c9dda115d63f8e3b38184391