Pith. sign in

Paper Citation Record · LEDGER

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2508.19578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19578 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:44:21.536125Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:42:38.614626Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40440c16-47f9-47de-9bae-06ba6c961d84 · outbound

This paper cites MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:21.329360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:44:21.329360Z digest=sha256:bf90b2c2b2f4236a377ca0755ec7134c05f442bda15e5363c7884b9c5b5bdf25

Observation e3a60b4e-3906-4055-9e88-ea59d07517c9 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:23.831621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.369285Z digest=sha256:bf3715b4e3b0ed6a1c77df82b60331d818895bf237bf511956e104112f2962b0

Observation d1e4b068-2c59-406a-9899-696486d65f83 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.187107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.363641Z digest=sha256:bdb4353cc7276921412510b45c499a4d05805a6717e5c85c65ddedc39f3f76cf

Observation 66d88c83-eb4c-44f1-b349-4fbcb40a7ac9 · outbound

This paper cites Output format: - Provide your answer in JSON format.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Output format: - Provide your answer in JSON format

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:23.379137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.382221Z digest=sha256:ca6459d06c9a3b7b71403692af81a61e3b296db5ab2e5599da3ac9061e66ee65

Observation 8b948acf-33f4-4e40-9972-0c54690b6c27 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:23.074028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.386736Z digest=sha256:c23e3387bd571ad871f2228c9e968ca22f8288b357d2477ed41e76a6dafc8128

Observation 8e9a3a28-a2b0-4f95-b0e9-69652deae086 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.484009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.395708Z digest=sha256:df78352aa45ed8439ad5dead2449784f5d4172b99ebb93b58249d703823e7f6c

Observation bd82aab6-d98f-495f-82a8-ca8902c217ed · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.136446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.410228Z digest=sha256:0c647080a99eb98fd3ca73f04038115f779f7238ce4dfebb7ba47082924e382e

Observation ebc4c111-a509-49dc-bf6e-4705aadf5765 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.753469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.427850Z digest=sha256:b9afbbf2db9c318222a7dabaad9881a5141ded73218ee10c0eea2ea7c0ce3a93

Observation ce833a26-02ff-4b1a-95e0-087abaa87f39 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:24.458041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.443220Z digest=sha256:a9a983967c8def8c59677d508cbd1a329f97afb48bf6ef4e966187ab6aee2768

Observation be50bfde-825f-4eb1-9876-84b2ea7b5fd2 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:22.864347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.463063Z digest=sha256:7ffe35ac8e1b9779dbdeb09107acf51a34a028a5eab8e97a1f46cc70e708ddfe

Observation 0dcda1b1-ff27-4801-bd7e-c463af71dcbe · outbound

This paper cites the protagonist.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts the protagonist

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:23.608046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.478356Z digest=sha256:78d1c2efa2f342cdf5459c1ea065ba2a2d21202025f2c8506cac4eb1980d5ebd

Observation 7c3e32ed-b6a9-46ec-97c1-5d71c21b9b41 · outbound

This paper cites sentence.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts sentence

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.612992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.494555Z digest=sha256:270074889b33f3108ba978af40ae08b71e821801c3f6de5b6a8c0382860a125c

Observation 779495c6-136d-4c84-90f8-5c56259ec96e · outbound

This paper cites • Less than 25 MFs → Score 1-2 • 25 - 35 MFs and at least a two-level hierarchy → Score 3-4 • More than 35 MFs and a well-defined multi-level hierarchy → Score 5.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts • Less than 25 MFs → Score 1-2 • 25 - 35 MFs and at least a two-level hierarchy → Score 3-4 • More than 35 MFs and a well-defined multi-level hierarchy → Score 5

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.354772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.509099Z digest=sha256:2a8a78b3c1665292995aef99df56855ad0e9a08a7fdd914f2bce0623fddb3db5

Observation 5b3408cc-8787-4944-9bd6-27547e8d724a · outbound

This paper cites 1-2 / 1" range • 3 points= A = YES but exactly one of B-D in the.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts 1-2 / 1" range • 3 points= A = YES but exactly one of B-D in the

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:22.044655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.521954Z digest=sha256:2e951b57ce55db39dd4f3f503fd455ca20f270deb8814b39866d291df08fb483

Observation 787d2cf1-3b53-4bdc-8ff6-e94ef252dbcb · outbound

This paper cites HKF_validity.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts HKF_validity

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:44:21.758590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.536125Z digest=sha256:74f53e4240f7f7c69dbf4b8495bda9e58b3b960a083c6a6b3a4fb83ca9595dc1

Observation 6b9d5de5-5b61-4411-aa16-363f63b34525 · outbound

This paper cites Is ChatGPT a Good NLG Evaluator? A Preliminary Study.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Is ChatGPT a Good NLG Evaluator? A Preliminary Study

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T15:44:21.334991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:44:21.334991Z digest=sha256:a6984dfec9bb21590c41fce027f0385f8edce98f5f85f55321710ad314afce2d

Observation 027bfa4d-8f02-4229-b15b-89342c7865a1 · outbound

This paper cites an unresolved cited work.

Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:44:25.774702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T15:44:21.323255Z digest=sha256:dab18af3e6a1ba0c09b870046fc4c84fbd29ebfcb352e0dca3d36af1503d9eee

Pith citing papers

Observation 013e6555-b428-4ea6-90f3-0a00f04d1e91 · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs Towards a Holistic and Automated Evaluation Framework for Multi-Level Comprehension of LLMs in Book-Length Contexts

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:38.614626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:38.614626Z digest=sha256:941a90d300951d713c4057070a182d4a68aa9657f567e1142c123b254dbbcd0c