Pith. sign in

Paper Citation Record · LEDGER

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports

As of 11 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2505.17265.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17265 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:50:51.824352Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-19T04:35:43.872967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T04:37:04.005940Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eeccb613-3fee-49b0-8200-a1b9b422d97c · outbound

This paper cites Large Language Models are Few-Shot Clinical Information Extractors.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Large Language Models are Few-Shot Clinical Information Extractors

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:50.831728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:50.831728Z digest=sha256:5d5d9d3c70f0085b66cefcdacf582f7955fef24faa8c43a1399e3068fec15ad7

Observation 24c37b6d-71e6-46a7-93b6-790d8a5d6860 · outbound

This paper cites The Llama 3 Herd of Models.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.029148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.029148Z digest=sha256:e416960984ba224a46a7c844bd20e96c9e166cfebd06781d28595862003488d9

Observation 0cdae921-87ec-465a-85f5-9481d07a68f9 · outbound

This paper cites Biocreative v cdr task corpus: a resource for chemical disease relation extraction.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Biocreative v cdr task corpus: a resource for chemical disease relation extraction

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:50:52.633353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:51.273496Z digest=sha256:341f75ef5cd69177431db55562bd22a706f1f4cabe5ec4a32994093da339bddd

Observation f02de11f-5fe2-49c2-9da7-a16e824145cc · outbound

This paper cites Health- prompt: a zero-shot learning paradigm for clinical natural language processing.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Health- prompt: a zero-shot learning paradigm for clinical natural language processing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:50:52.486425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:51.643986Z digest=sha256:6802f638618fe25b5335e24e32ffc2c20cf001d2b618bda99b828bbc389303a6

Observation 6816fcb9-825c-4203-8634-fe5441c9d8d1 · outbound

This paper cites PMC-Patients: A Large-scale Dataset of Patient Summaries and Relations for Benchmarking Retrieval-based Clinical Decision Support Systems.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports PMC-Patients: A Large-scale Dataset of Patient Summaries and Relations for Benchmarking Retrieval-based Clinical Decision Support Systems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.824352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.824352Z digest=sha256:ae24bab1193d7ff0306005d3e884e8d2770761c918ce1cda21aaf5902597acdb

Observation 262a9bd2-3f3c-469b-93e5-f0f55ae34147 · outbound

This paper cites Measuring annotator agreement generally across complex structured, multi-object, and free-text an- notation tasks.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Measuring annotator agreement generally across complex structured, multi-object, and free-text an- notation tasks

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:50:52.966850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:50.961153Z digest=sha256:fdc495ccef211ab545897713dd7110a7b42c71eef86d4ede2386667da5bda0bd

Observation d2f07db3-4b71-4ae4-95bd-9491a456aade · outbound

This paper cites Lessons from Natural Language Inference in the Clinical Domain.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Lessons from Natural Language Inference in the Clinical Domain

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.556151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.556151Z digest=sha256:7822579787f3ae36c239bc4ba4d7344241ab8e81d6a98a09bcabecc7ee5e57b1

Observation 5825569e-c9e1-4aa8-b400-3175fd241188 · outbound

This paper cites Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting

Reference 2018

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:50:52.228891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:51.356196Z digest=sha256:8cdd386e0b91b9b99b0a95eabd95258beaa40cf978c4e8fd3a63484cc82563a4

Observation c0940e44-7b0b-40fd-b772-bdd914b1ca86 · outbound

This paper cites Medical information ex- traction with large language models.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Medical information ex- traction with large language models

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:50:52.821744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:51.104640Z digest=sha256:da88eb5a4569a0e88855b5a339a06841020676774c6c8b0be150ef0eb875beba

Observation deb4af2b-81dc-489d-8cde-fce44696cefa · outbound

This paper cites an unresolved cited work.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Unresolved cited work

Reference 2022

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:50:53.145420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:50:50.876720Z digest=sha256:33f60f924a11a8a813d2eba3382e998cfae1f3791e854fd55d011176c78d2aaf

Observation ef2f3c31-7ebb-4982-9a0c-b7329d726c1f · outbound

This paper cites Qwen2.5-Coder Technical Report.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Qwen2.5-Coder Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.193980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.193980Z digest=sha256:da5dbf29cebdf35ddde56f4177def3d8f4aa46be29af9aa4f94c2c16c8267b35

Observation 19e355af-07b8-4224-9b2b-6f5a8e6b7914 · outbound

This paper cites Qwen2.5 Technical Report.

CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports Qwen2.5 Technical Report

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.732581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.732581Z digest=sha256:fd1cef8698cfcf4be36a61ec5c3dcc88491f45185869c582d27eb51c7557e42c

Pith citing papers

Observation 9b9f76dd-6a5f-4103-9182-44f59c2dbfd8 · inbound

Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes cites this paper.

Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes CaseReportBench: An LLM Benchmark Dataset for Dense Information Extraction in Clinical Case Reports

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:37:04.008415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-19T04:35:43.872967Z digest=sha256:5c3bc803dec89e613577c61f70ac94d523708e016f4e7fe3fb6543feb1975fe1