Pith. sign in

Paper Citation Record · LEDGER

Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2402.04614.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04614 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:43:26.331690Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

11
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1828b972-a9dd-4c55-a2b8-1ecbf58bf901 · inbound

A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification cites this paper.

A Systematic Comparison between Extractive Self-Explanations and Human Rationales in Text Classification Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:48:22.968721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T19:48:20.738920Z digest=sha256:d0bd7c71e1940995c3efc21776f04131a3b5536e39e44366e062861375d105a6

Observation b9f1e6d8-dbaf-4d73-9b7b-12fc3a257cbc · inbound

OpenAI o1 System Card cites this paper.

OpenAI o1 System Card Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:42:39.633867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T06:39:44.542350Z digest=sha256:994668eb1e877b2458295922915bfb1225b9c1cb7b15baf28528fa0faae618cd

Observation 0a5f5c02-66dc-4a2a-9f98-10b77a9b0c00 · inbound

Fostering Appropriate Reliance on Large Language Models: The Role of Explanations, Sources, and Inconsistencies cites this paper.

Fostering Appropriate Reliance on Large Language Models: The Role of Explanations, Sources, and Inconsistencies Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T04:43:26.331690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:43:26.331690Z digest=sha256:56aa81b152ea870583fa2d92c86f8320e612c390ce43cc382a9b13becff6f971

Observation f2478e21-7dc4-4a0f-ac57-57f1501d7302 · inbound

PEDANTIC: A Dataset for the Automatic Examination of Definiteness in Patent Claims cites this paper.

PEDANTIC: A Dataset for the Automatic Examination of Definiteness in Patent Claims Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:52.285363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:52.285363Z digest=sha256:df94a33b86a8fc73d92a6d00c5deba3839e56ea84f708f5413fb840d5841fbd1

Observation 144b3316-b6ca-4b62-aaea-a5f115d28237 · inbound

Energy-Based Transformers are Scalable Learners and Thinkers cites this paper.

Energy-Based Transformers are Scalable Learners and Thinkers Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-06T20:42:39.003678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:42:39.003678Z digest=sha256:02e5ef8c19c78547dc3c417b84ce421c5163bb05f23474d44378b5fc901abb83

Observation d98a4142-9ed6-4965-b283-9bdba0e44275 · inbound

SynthEHR-Eviction: Enhancing Eviction SDoH Detection with LLM-Augmented Synthetic EHR Data cites this paper.

SynthEHR-Eviction: Enhancing Eviction SDoH Detection with LLM-Augmented Synthetic EHR Data Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T18:47:53.651283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:47:53.651283Z digest=sha256:a9088edcd48f7443b744a1e115139fd4a8bdda227c709b74a35a556d8b9141b2

Observation 5ce31e31-5653-48e1-87c7-f75340536d87 · inbound

Neither Valid nor Reliable? Investigating the Use of LLMs as Judges cites this paper.

Neither Valid nor Reliable? Investigating the Use of LLMs as Judges Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T16:40:06.296443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:40:06.296443Z digest=sha256:e00aa995dd5ea6dce715e1fd7e9daf70e39de03eb768eb7e3a66604bc8d11e97

Observation 3a4e82b5-efdf-4f41-8d7f-0ffc67117d05 · inbound

Explainable Knowledge Graph Retrieval-Augmented Generation (KG-RAG) with KG-SMILE cites this paper.

Explainable Knowledge Graph Retrieval-Augmented Generation (KG-RAG) with KG-SMILE Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-05T10:53:13.491095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:53:13.491095Z digest=sha256:2fc992714b54aa0d47726dd09de4ee517dcdb496f4079717bb99ae829c295bec

Observation 3f73ef14-9207-4f9d-bf08-4e403d033815 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:04.627878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:04.627878Z digest=sha256:e2fde1407a8b90a5d91d6840c367ed3d92b360300ac63521be4819cef5b0ea36

Observation 495e8aac-0b32-4581-a4d9-f06d8bb96e73 · inbound

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models cites this paper.

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:04:52.473308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T13:04:42.614541Z digest=sha256:ff17ab45082d5fcbb00842da2fc2606c1c4fd9ae6b41f2e8da0220e908e48486

Observation bfa6dc8f-ec8c-4cc0-85fa-7d0e92de91c7 · inbound

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models cites this paper.

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T08:29:33.592184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:29:33.592184Z digest=sha256:a773810c9a81c9d84c95ae66d9d99002258a0150ae6ba474bf1003504e253c63

Observation c4aef10c-03fa-45b5-9e66-e7cdc91b6fb5 · inbound

When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making cites this paper.

When AI Persuades: Adversarial Explanation Attacks on Human Trust in AI-Assisted Decision Making Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:24:11.351546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T13:21:45.525765Z digest=sha256:c000bf90ac325c4e83004d42a2113ee5d277f53aba9f75a71b6d8e1dfc0aba9a

Observation 420c0c12-68bc-48e6-b90a-599b31d4ad00 · inbound

Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation cites this paper.

Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generation Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T23:17:13.981535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:17:13.981535Z digest=sha256:2231be16566f5b0c5dfcc73168c9e7894cdee9699eebb8e049e31b9629f7968b

Observation 634ec39a-2a5b-44ed-9e38-0301e73496dc · inbound

TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning cites this paper.

TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T21:53:25.091493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:53:25.091493Z digest=sha256:b696e7843e353202170fc1898bbe8db844d851f6266f5f1998e43b245ed5b4d2

Observation 65ddf80e-8132-4058-b452-d9b35916ccaa · inbound

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't cites this paper.

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:00:49.734036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:24:48.212344Z digest=sha256:7e02a4da87c83d2a67dc71e4fd46cbfbf91fa32bf56f0300f8b965ef448defbc

Observation 16587e41-b78f-4818-8669-a7f71ed1049e · inbound

AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliency cites this paper.

AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliency Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.731707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T08:08:50.330857Z digest=sha256:db42015ae8a123b2d17f1e18ee6bf181f515a2992a87e4014a1b3948721e7412

Observation 441a134a-cd48-469c-891b-c2f04d6bc9df · inbound

PageGuide: Browser extension to assist users in navigating a webpage and locating information cites this paper.

PageGuide: Browser extension to assist users in navigating a webpage and locating information Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:26:13.928728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T05:39:04.895180Z digest=sha256:6def7921abd2227fcc7605f2a01ff4c5c8087841537213d249488f1348cdf7e7

Observation 797f06fd-cb6e-4c8e-9262-6e20199469ef · inbound

PageGuide: Browser extension to assist users in navigating a webpage and locating information cites this paper.

PageGuide: Browser extension to assist users in navigating a webpage and locating information Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:25:40.413243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T09:08:44.608520Z digest=sha256:fe824211f92399f0a3768c9acb3a5aaa86c39d518dac85db58daba2205076a9a

Observation 29f32337-e190-4ecd-933e-08b05ba10e5f · inbound

Concept-Based Abductive and Contrastive Explanations for Behaviors of Vision Models cites this paper.

Concept-Based Abductive and Contrastive Explanations for Behaviors of Vision Models Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:39:25.070594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T12:05:51.787664Z digest=sha256:78a3aef92b2cfc10e67f5a74e03d4bd48de4cd075141413e9a00e0d44d9ecefe

Observation 503ffb3c-ed7e-470f-a720-d5bb57902e48 · inbound

Interpretability Can Be Actionable cites this paper.

Interpretability Can Be Actionable Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 132

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:17:23.259684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T06:12:51.656452Z digest=sha256:bc7d7fcb0ebab4c361946596c6d876e85904e49c4add556f652611ed8e76e97c

Observation 5218c4cf-44c0-4dd9-bfcf-44168abb726f · inbound

Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization cites this paper.

Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:24:39.914596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T12:17:12.602012Z digest=sha256:a205f02d9ac01f0a09cf8af88afd2290322a7c68a868c96efe3ff1497cd36bf3

Observation b2c7a080-ba62-464e-b328-97a75077818a · inbound

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth cites this paper.

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:04:39.040687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:56:53.355299Z digest=sha256:08ed216b12fa6a6889003a73be576c15aa1925d3bff3fe5c87fca1534881b1c0

Observation 44708cbb-810d-4df2-8205-f004085547ab · inbound

Detecting Unfaithful Chain-of-Thought via Circuit-Guided Internal-External Discrepancy cites this paper.

Detecting Unfaithful Chain-of-Thought via Circuit-Guided Internal-External Discrepancy Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:53:59.506942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T21:47:17.894881Z digest=sha256:3f8589d6233fd84837873f3a33e4b6a6c2b9e91068ca254e59dff86d8b37e657

Observation 35682200-1dbf-4f97-ab2e-a560cc459951 · inbound

A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales cites this paper.

A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.819616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T16:33:41.230574Z digest=sha256:c9697f999a4ebfd0a90910a9cabd2e166bf008732dc64db3f0f84e9099d3b43c

Observation f01eac34-ad72-47d9-bc7c-7704f0dc265a · inbound

Forecasting Future Behavior as a Learning Task cites this paper.

Forecasting Future Behavior as a Learning Task Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-03T06:07:41.352435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T12:55:39.494339Z digest=sha256:2933d3fb6bf2a17dbf2cd53ee8f9acf3e1907f9b105541a05ea85e2ccc074e43

Observation bd78a7e4-a287-4a31-9cf6-02b8c4254a3d · inbound

Answer Engineering: Local Trajectory Editing for Protocol-Constrained Decision Making in Large Language Models cites this paper.

Answer Engineering: Local Trajectory Editing for Protocol-Constrained Decision Making in Large Language Models Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:49:37.567471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T14:13:13.678245Z digest=sha256:9433bda1024fed96e2f31aee163cdff65306496fa120490123b22505c09d8e03

Observation 4a643261-618e-4dc1-9e7f-27e38fed8d25 · inbound

BetXplain: An Explanation-Annotated Dataset for Detecting Manipulative Betting Advertisements on Social Media cites this paper.

BetXplain: An Explanation-Annotated Dataset for Detecting Manipulative Betting Advertisements on Social Media Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 286

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:19:50.813832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T05:18:36.311573Z digest=sha256:fd394c02fced5737adab0a2dd3732deb724e7ccc20b242a41305d4a183f57ff5

Observation 3175cf0c-4abd-4829-8573-f576a4499399 · inbound

What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs cites this paper.

What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefs Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:25:50.071794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-30T00:38:21.949283Z digest=sha256:6d45b68505d8a4b2c0d8ac9d2de425275e87ff682151cd489472b495f6ba6121

Observation 2f497e06-a006-48af-b1b9-13c856bc6047 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.481838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:a3cdf75cf78dcbfec8b1bc2164cc20210e0e15339517ce6fdef1ed806031494c

Observation e90c452e-82f0-4b0b-a981-db1bbef2ef76 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.698322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:1c86787aa7082d27b207fe72cdf68de855ab33ce1d0981f7cf0c9525c4377682

Observation 01ed0576-4136-4793-8374-48cc2f9be613 · inbound

Scientific Explanations in Health Sciences: Causality, Trust, and Epistemic Adequacy cites this paper.

Scientific Explanations in Health Sciences: Causality, Trust, and Epistemic Adequacy Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:35:41.485965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:24:21.103206Z digest=sha256:6632c8f9454433274fd1a8a8008c824c34ff9e4190b4560c1764a7e9d6f48bf5

Observation 588ccba3-18bc-4e6f-bb77-51e056a72c74 · inbound

Governing Generative AI Across Financial Institutions: A Framework for Generative AI Risk Control cites this paper.

Governing Generative AI Across Financial Institutions: A Framework for Generative AI Risk Control Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T21:40:45.262610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T21:40:45.262610Z digest=sha256:91f48126fdce27d64301c63753367059392cdbb06be5ad69815aa4dbadcfeba6

Observation 15b68ccd-3771-4e59-bc9e-ace8816dba30 · inbound

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins cites this paper.

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T07:44:08.123559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:44:08.123559Z digest=sha256:6cb32806e4b735a2e717c4034c422543686e5760504416cb8575e67e7bfa1c54

Observation 4fdba0a0-5c6f-4cc9-b3ad-c2fa419090fb · inbound

From Plausible to Actionable: A Position on LLM Self-Explanations cites this paper.

From Plausible to Actionable: A Position on LLM Self-Explanations Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T21:49:33.224252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:49:33.224252Z digest=sha256:2226b41f96505fe9147b029df2f2ad794ec7612ba53f0135f88354ece767fff7

Observation 479a0402-854e-4705-87ba-8af21c455c77 · inbound

Mechanistic Attention Guidance for Agent Memory Refinement cites this paper.

Mechanistic Attention Guidance for Agent Memory Refinement Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T17:30:38.471354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:30:38.471354Z digest=sha256:31ef84a2ff6888aba4c9ecc6703284f8400d306f1ce9581807308460d0472ca8

Observation 773c65b4-8020-4f03-a4a2-bfc25386ce49 · inbound

Training Large Language Models for Self-Explanation Faithfulness cites this paper.

Training Large Language Models for Self-Explanation Faithfulness Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T08:36:18.674676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:36:18.674676Z digest=sha256:5696232b75bfa82cf921b4fcbedafc9f4b0764c1bdbca1ef760d82843265daad

Observation 304a171f-1d9c-4c77-a715-fe5e3471a408 · inbound

To Facilitate or not to Facilitate: Human and LLM Facilitator Tendencies in Online Discussions cites this paper.

To Facilitate or not to Facilitate: Human and LLM Facilitator Tendencies in Online Discussions Faithfulness vs. Plausibility: On the (Un)Reliability of Explanations from Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T00:53:08.096173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:53:08.096173Z digest=sha256:ca4cf9f86e3a5bde6f46cce68ab3b24fd5bc2444979cd9be275260f385ba8e93