Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2403.03558.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.03558 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:46.855329Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T10:35:42.145281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 871f0a7b-47a7-47b6-934d-96c044af44fb · inbound

Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions cites this paper.

Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 168

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:31.399061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:31.399061Z digest=sha256:6dacd712733d6ac64ed41c2077d5f6a27a0421c41d15735dc1b70a19758c6ae4

Observation ec71614e-0c23-4d31-9eb4-2a39d21b4b7c · inbound

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models cites this paper.

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:18:46.855329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:18:46.855329Z digest=sha256:d77422ea80e8c107fdf358573ea3fd703e45bd09096a06213380f8ecd2f9dfb5

Observation 4047f972-478e-4ed5-894c-c816aaaf8507 · inbound

VisionTrap: Unanswerable Questions On Visual Data cites this paper.

VisionTrap: Unanswerable Questions On Visual Data Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T14:57:08.303386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:57:08.303386Z digest=sha256:13aa30e2539617c7e7fd563402b261c74d8522dc1276ec354b9de46db49af8fb

Observation f618e602-033b-485e-8bfe-c48d9bc60dae · inbound

Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA cites this paper.

Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T22:18:43.463462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:18:43.463462Z digest=sha256:b8243fe19db14e3252d22f33c52f9d6e082710f8328fb13cb1729a187d7a8e4e

Observation 30fd6803-76c7-4009-9a5d-4c3c208ccb52 · inbound

Neural Message-Passing on Attention Graphs for Hallucination Detection cites this paper.

Neural Message-Passing on Attention Graphs for Hallucination Detection Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:12.288648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:12.288648Z digest=sha256:3f5b32fdf1006cec81d7a24a9eb5c4df1543cf5364817564523b74475d9ababa

Observation 9aca73f5-75ae-4ec9-b117-543e1dce8d0d · inbound

Enhancing LLM Metacognition via Cognitive Pairwise Training cites this paper.

Enhancing LLM Metacognition via Cognitive Pairwise Training Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.895497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:aaad97793233f1e75cc60775db52d47cd3d9fcab68131e6834c8c8522d1b23b0

Observation 08766ef6-6be5-4466-9de4-2433c71da695 · inbound

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs cites this paper.

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:35:42.147094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:22:38.232552Z digest=sha256:11abfc1b8976340497e6968e9f5c11a25cbc499d93556ef078eafce4c95cfaff