Pith. sign in

Paper Citation Record · LEDGER

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts

As of 23 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2508.19578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19578 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:44:21.536125Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:42:38.614626Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40440c16-47f9-47de-9bae-06ba6c961d84 · outbound

This paper cites MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:21.329360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:44:21.329360Z digest=sha256:9368bdb18ebd1b509367d3ed8fecec0b1e5c0af62a864d240af68df29f0c3e10

Observation e3a60b4e-3906-4055-9e88-ea59d07517c9 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:23.831621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.369285Z digest=sha256:3ef0cc503a2b85583e38a6ccb1079397443748ed607cac14e81bfd65c17169ba

Observation d1e4b068-2c59-406a-9899-696486d65f83 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.187107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.363641Z digest=sha256:227a9ea5484f64f6764052f35906a17d98bc9d91941265e8d5213e1182e50582

Observation 66d88c83-eb4c-44f1-b349-4fbcb40a7ac9 · outbound

This paper cites Output format: - Provide your answer in JSON format.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Output format: - Provide your answer in JSON format

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:23.379137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.382221Z digest=sha256:7dc52788f702149786419a88de1058922cc6b7f8b0d767b959a1e1f14b03f7d6

Observation 8b948acf-33f4-4e40-9972-0c54690b6c27 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:23.074028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.386736Z digest=sha256:7441a18a22e7bb3a2121e1e404aef87c031572411425d8f89e5c61b7d407385a

Observation 8e9a3a28-a2b0-4f95-b0e9-69652deae086 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.484009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.395708Z digest=sha256:85f98d2c9305d460d561339829b9ceacdbf284c6bd5fc544637afcf4c58a6f63

Observation bd82aab6-d98f-495f-82a8-ca8902c217ed · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.136446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.410228Z digest=sha256:3766557dd8d3aee3c05be32f35f95093ec0bf244f6a264326caded78d31468f3

Observation ebc4c111-a509-49dc-bf6e-4705aadf5765 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.753469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.427850Z digest=sha256:78b342bdfe5e13c98392e077b9545f45c3de4c0d7b64b5f7573672a95e03b18f

Observation ce833a26-02ff-4b1a-95e0-087abaa87f39 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.458041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.443220Z digest=sha256:78affef0a6e8c1dce0ae0c96d971b14a8b3cff44bf48ba2b14719257477cfe88

Observation be50bfde-825f-4eb1-9876-84b2ea7b5fd2 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:22.864347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.463063Z digest=sha256:e1545da1c1fcbf361e34820000fee4f214f2f26ddb31763c42c64d7e43ebdfa7

Observation 0dcda1b1-ff27-4801-bd7e-c463af71dcbe · outbound

This paper cites the protagonist.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts the protagonist

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:23.608046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.478356Z digest=sha256:8c100bcdbc01362bbe2cf368d0378f46a284237a0c8ed3d186abdef4c568a607

Observation 7c3e32ed-b6a9-46ec-97c1-5d71c21b9b41 · outbound

This paper cites sentence.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts sentence

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.612992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.494555Z digest=sha256:ff521867c39d4d397fb07fea705f13f2f1a5aa721a0189694222c34eb170ec34

Observation 779495c6-136d-4c84-90f8-5c56259ec96e · outbound

This paper cites • Less than 25 MFs → Score 1-2 • 25 - 35 MFs and at least a two-level hierarchy → Score 3-4 • More than 35 MFs and a well-defined multi-level hierarchy → Score 5.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts • Less than 25 MFs → Score 1-2 • 25 - 35 MFs and at least a two-level hierarchy → Score 3-4 • More than 35 MFs and a well-defined multi-level hierarchy → Score 5

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.354772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.509099Z digest=sha256:d615e42837e14d05e630d8bc36aee9ceb894cd372e8f3c876538429bc8d4a8ed

Observation 5b3408cc-8787-4944-9bd6-27547e8d724a · outbound

This paper cites 1-2 / 1" range • 3 points= A = YES but exactly one of B-D in the.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts 1-2 / 1" range • 3 points= A = YES but exactly one of B-D in the

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.044655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.521954Z digest=sha256:d57c8a722248497d554429b9f280d4be51104e9b988b8fb0899f77e285efff89

Observation 787d2cf1-3b53-4bdc-8ff6-e94ef252dbcb · outbound

This paper cites HKF_validity.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts HKF_validity

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:21.758590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.536125Z digest=sha256:60953d961a6788fa1a926dc650a046a71fd8442a944a06d64398f598e47ed116

Observation 6b9d5de5-5b61-4411-aa16-363f63b34525 · outbound

This paper cites Is ChatGPT a Good NLG Evaluator? A Preliminary Study.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Is ChatGPT a Good NLG Evaluator? A Preliminary Study

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:21.334991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:44:21.334991Z digest=sha256:1da212e8d73142065da8968e4de789b7021d0b380fb87a4cd976e85bb094b3c4

Observation 027bfa4d-8f02-4229-b15b-89342c7865a1 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.774702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T15:44:21.323255Z digest=sha256:9b05a438cd74e8ce4c262095f156d9f16311d52a051245f6046d04525ddd9dda

Pith citing papers

Observation 013e6555-b428-4ea6-90f3-0a00f04d1e91 · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:38.614626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:38.614626Z digest=sha256:6005e3db352edb3f2a8c96ae7dafcfff0282d78396b65bbadde9174ac35a8ee7