Pith. sign in

Paper Citation Record · LEDGER

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models

As of 17 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2606.05170.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.05170 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T20:40:22.043187Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a4911c00-4557-4639-b149-41d5ba070bb7 · outbound

This paper cites Beyond accuracy: Measuring the severity of LLM hallucinations.Findings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP Findings),.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Beyond accuracy: Measuring the severity of LLM hallucinations.Findings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP Findings),

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:dd04247d02bbc25919919d85c87ca6fd15628ad240bac912f5e47b1c4b68a2af

Observation 60fddd3c-71ac-4c75-9e29-a060c8b301ab · outbound

This paper cites Lost in tran- scription, found in distribution shift: Demystifying hallucination in speech foundation models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Lost in tran- scription, found in distribution shift: Demystifying hallucination in speech foundation models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:8be39ee3a06c0a68c091ea3491d79bf407f3fcb3e7dc7ab109b7d5277879eb7a

Observation 2ea54fda-b07b-457e-8118-6a44826b5be9 · outbound

This paper cites MedHEval: Benchmarking Hallucinations and Mitigation Strategies in Medical Large Vision-Language Models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models MedHEval: Benchmarking Hallucinations and Mitigation Strategies in Medical Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:80ecd31ffad06057b68fa0296130b815093d5a4ae51b1d948676ac3bb438e800

Observation a3263dd9-704d-4909-ac4d-52917315e976 · outbound

This paper cites Overview of the ClinIQLink 2025 shared task on medical question-answering.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Overview of the ClinIQLink 2025 shared task on medical question-answering

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:5ad79b798defe25d64311488c0b3cb9a56a7e32ebd6aa2bce82b835170cf842c

Observation 4b3e4af2-0ca1-494b-b2e5-83225a8d7edc · outbound

This paper cites an unresolved cited work.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:944fc10972f5ba0e52f4d7acfde39cbd1597f139c9274f3957c50b372951f2a3

Observation 19380e45-05b4-4af7-9b00-c4063181e08a · outbound

This paper cites Language Models (Mostly) Know What They Know.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Language Models (Mostly) Know What They Know

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:ae11771f504078a58f7f471cb72fb8e801f4aa0d24c583d0da16c8f15975cab3

Observation 4b37c6ea-ff45-431e-9613-97fc45cb776a · outbound

This paper cites Scaling Laws for Neural Language Models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Scaling Laws for Neural Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:8fc9d711c144bf64ccea82b7e1eb8d7de2fe9bdc0d025e4ff3c8e823107c3453

Observation ee516dd6-88b7-41bf-80ea-00bb804b0177 · outbound

This paper cites HaluEval: A large- scale hallucination evaluation benchmark for large language models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models HaluEval: A large- scale hallucination evaluation benchmark for large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:23b7835bb4b8d13c7f2c495036351cad829050fd73017cab58534c0501446fd6

Observation cd9bb82f-2779-4149-b9bb-b86d719790f6 · outbound

This paper cites Teaching models to express their uncertainty in words.Transactions on Machine Learning Research, 2022a.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Teaching models to express their uncertainty in words.Transactions on Machine Learning Research, 2022a

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:9561b7d65ec9e5af9def9c71533533829fd38971c5969fa3ba80b060d73b89d7

Observation 1bc3ecaf-75e3-4dcb-a42f-b1443d67e200 · outbound

This paper cites Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models Towards a Systematic Evaluation of Hallucinations in Large-Vision Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:f4aeb5fb70d60db589bb76c751aad852c10d003a15eef98584530743a694696f

Observation 1e92785f-c4af-4572-bda8-1688b7c71bfa · outbound

This paper cites MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:6f233a47d4f315b2532d7025ef7ef6b3a2e3298cbdd9966e178b78b4aa663c0b

Observation 73086346-b225-46c2-a7e4-180a95dcdc40 · outbound

This paper cites verbose hedge.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models verbose hedge

Reference 12

Resolution
malformed identifier
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:c404f73092bde7cacb4e8bcd3d00d9e3eda3cdfbce45d195a41416b60934d885

Pith citing papers

No inbound Pith citation observations are available.