Pith. sign in

Paper Citation Record · LEDGER

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents

As of 14 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 3 inbound Pith citation observations for arXiv:2605.06635.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.06635 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T09:53:00.677605Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:27:14.690819Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-05-21T02:13:56.467650Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact8
  • verified fuzzy1
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc9467c9-71cd-48b7-bd66-ab6a38d8b1e0 · outbound

This paper cites CiteGuard: Faithful Citation Attribution for LLMs via Retrieval-Augmented Validation.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents CiteGuard: Faithful Citation Attribution for LLMs via Retrieval-Augmented Validation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T20:16:08.686759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:d53b331820ab727640db689880c604e392274686fcfa0abe02803855f13b753e

Observation c39df063-a5fd-43f5-accb-7a6bddd64b07 · outbound

This paper cites DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:07:39.643626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:21957de8c11a76d582b4f83c10e645dbca1733de1c7d96cc647f4d2299be4869

Observation 935be28c-0af6-42fe-9a3d-f6b720f4dade · outbound

This paper cites RARR: Re- searching and revising what language models say, using language models.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents RARR: Re- searching and revising what language models say, using language models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.720550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:68e900dccea821fc44cbdad4f8c3f44551f6138294130f4519d582ea2816c6f4

Observation 0d45735f-5c26-415b-98f3-115398b35a70 · outbound

This paper cites AttributionBench: How hard is automatic attribution evaluation? InFindings of the Association for Computational Linguistics: ACL 2024.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents AttributionBench: How hard is automatic attribution evaluation? InFindings of the Association for Computational Linguistics: ACL 2024

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T15:42:36.591406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:e27f54e7d3e09b19008cf5e9d412314929faa204b808f4fef57d959d8a1ef62c

Observation 9e019c9c-7940-4b2d-acdb-9a104aba70fc · outbound

This paper cites Comparison of text-based and image-based retrieval in multimodal retrieval augmented generation large language model systems.arXiv preprint arXiv:2511.16654, 2025a.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Comparison of text-based and image-based retrieval in multimodal retrieval augmented generation large language model systems.arXiv preprint arXiv:2511.16654, 2025a

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.675571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:56115559b5514540a1d070eff0d52c63a301291ea573aebed1d84e74c498f61d

Observation f012783a-b88d-4672-a60a-6680bf776866 · outbound

This paper cites Burke , title =.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Burke , title =

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-08T22:14:17.380945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:e42f299a6a81a575b1bb311c2a0a02748366f695e1174b8f9235e5f8bf653546

Observation 5d441d37-26ec-4dc1-8b4f-bff7afeb7823 · outbound

This paper cites HALoGEN: Fantastic LLM Hallucinations and Where to Find Them.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents HALoGEN: Fantastic LLM Hallucinations and Where to Find Them

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:16:08.740692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:8c302debdac0c2cc094f1809a4d444fa1b58c3d981583cc60ef19461cb94d681

Observation 3063c321-49b5-4d90-9c10-5a684d552aaa · outbound

This paper cites Yash Saxena, Raviteja Bommireddy, Ankur Padia, and Manas Gaur.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Yash Saxena, Raviteja Bommireddy, Ankur Padia, and Manas Gaur

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.701686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:7066eedc72f79a7d3cbabc7780fee0c3e4ba8513252fbec56027fd6bc47f0453

Observation ec0b2f4a-02b8-40a8-a9d7-89c21d1e2561 · outbound

This paper cites Chronos: Temporal- aware conversational agents with structured event retrieval for long-term memory.arXiv preprint arXiv:2603.16862.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Chronos: Temporal- aware conversational agents with structured event retrieval for long-term memory.arXiv preprint arXiv:2603.16862

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.734809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:8ff894d2124e3a56e5a9abab056bcaff006d40a0852ee9353002e29831af1fa5

Observation 620b54f3-9ecf-4f28-8392-a6ec753257bc · outbound

This paper cites Verifying the verifiers: Unveiling pitfalls and potentials in fact verifiers.arXiv preprint arXiv:2506.13342.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Verifying the verifiers: Unveiling pitfalls and potentials in fact verifiers.arXiv preprint arXiv:2506.13342

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.750767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:dba6cb57d9e3d5ed0f611521d158324f03de2ebd020742cbb57125eb77197394

Observation 4e3dc6ab-b490-4765-abf3-4665511a40d6 · outbound

This paper cites GenerationPrograms: Fine-grained Attribution with Executable Programs.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents GenerationPrograms: Fine-grained Attribution with Executable Programs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.692545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:45b9a3f8ebc39935031b6a7b32016bb68a4a672b9837d5169e2cb830e8f67515

Observation 04ef2b08-c858-4f54-97a1-7758616c055c · outbound

This paper cites Assessing Judging Bias in Large Reasoning Models: An Empirical Study.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents Assessing Judging Bias in Large Reasoning Models: An Empirical Study

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:08.764776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:1aae46b66c21ed4c67fcf0099a2c0569fce859b777cb766401bbcd91f024f218

Observation 970abbb4-88e3-4ec4-a4a4-edb8b8d6a29e · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:44:32.193354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:313e6d77c7b5c6b246f2f94f49bab9afa70044517c2845b9ca96c3b29c640980

Observation 0e1285aa-ec53-43a0-be95-9d1c8bf57c4d · outbound

This paper cites CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era.

Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents CiteAudit: You Cited It, But Did You Read It? A Benchmark for Verifying Scientific References in the LLM Era

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T20:16:08.728420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:53:00.677605Z digest=sha256:49aacf06afc42ec734214913912c8bbc41e80f9e464db3ccf2506f21087d9416

Pith citing papers

Observation 25717fdb-e762-4faf-a57f-c48967085b44 · inbound

Pramana: A Protocol-Layer Treatment of Claim Verification in Autonomous Agent Networks cites this paper.

Pramana: A Protocol-Layer Treatment of Claim Verification in Autonomous Agent Networks Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-21T02:13:56.470628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-21T02:10:13.030288Z digest=sha256:fdd6ac3f30b94565c84c4d5259d5323dea6b6b721a7f5084be1459d511285ab4

Observation 2eb4d342-87ea-40cf-aa69-1a54321e0a3e · inbound

VetScore: Risk-Weighted Fact Verification for Veterinary Long-Form QA with Citations cites this paper.

VetScore: Risk-Weighted Fact Verification for Veterinary Long-Form QA with Citations Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:02:59.655565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:02:59.655565Z digest=sha256:fb66dafa10bfc51a0a73d89923b7297cdff88e8f4514f6930f8a23f3a4263ebe

Observation c4fff4a4-17d5-4146-894e-98057756c811 · inbound

V-FiLLM: Verified Financial LLM Reasoning Benchmark cites this paper.

V-FiLLM: Verified Financial LLM Reasoning Benchmark Cited but Not Verified: Parsing and Evaluating Source Attribution in LLM Deep Research Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T11:27:14.690819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T11:27:14.690819Z digest=sha256:0efefe3946f2b271ba98baf66cca7f87eec30ff540dfd2940c7274210abd45ba