Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Large Language Models in Retrieval-Augmented Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2309.01431.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.01431 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:52:15.422851Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:59:56.182550Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 33845667-6395-494e-8b4e-83727b5c8f1e · inbound

Retrieval-Augmented Generation for Large Language Models: A Survey cites this paper.

Retrieval-Augmented Generation for Large Language Models: A Survey Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 167

Resolution
verified exact
arxiv_id, observed 2026-05-24T05:13:56.515375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-24T05:10:25.171044Z digest=sha256:2ac7878cef04643693502367ee30b6116b2eb160db7abcb65805f125e0b72e31

Observation e212d72d-dc8a-4fab-b5a0-8479cbd48790 · inbound

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries cites this paper.

MultiHop-RAG: Benchmarking Retrieval-Augmented Generation for Multi-Hop Queries Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:54:19.677200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T13:54:19.264893Z digest=sha256:1221078e63830708da01b326b35821198f1cdb0df1f7609972f78f823226f42e

Observation 52057d17-e24d-4106-9d4c-8b900e9c9d11 · inbound

Disrupt Your Research Using Generative AI Powered ScienceSage cites this paper.

Disrupt Your Research Using Generative AI Powered ScienceSage Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T00:52:15.422851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:52:15.422851Z digest=sha256:99b4f0a05547340fce85a76a8111f751382718b973e409249ea1ba1b94efea26

Observation ff75edf6-ce90-4e86-8e1d-97709f4fc640 · inbound

FinS-Pilot: A Benchmark for Online Financial RAG System cites this paper.

FinS-Pilot: A Benchmark for Online Financial RAG System Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:33.958225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:33.958225Z digest=sha256:f90e4b97d7cd253a495683e4ec54ad643887bee639d044a6f7a60723c405e83b

Observation 1c0023b4-c832-424b-a09d-244b1892ae4c · inbound

Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions cites this paper.

Evaluating and Improving Robustness in Large Language Models: A Survey and Future Directions Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:30.832377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:30.832377Z digest=sha256:eeda2daa35a95e6a0c1d385bb31f4a42cf62049b4b487a12115b81d8b485f264

Observation b07cf664-b701-4a39-a7aa-72b453789ac8 · inbound

HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation cites this paper.

HIRAG: Hierarchical-Thought Instruction-Tuning Retrieval-Augmented Generation Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:24:50.354323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:24:50.354323Z digest=sha256:0fc8df5067a27e05afd58a68928acc2c8a14631649b6d7b54e2f375c1e822929

Observation 14f09286-b735-4ce0-8edd-f3ddb910014e · inbound

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level cites this paper.

Investigating the Robustness of Retrieval-Augmented Generation at the Query Level Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:55:54.682167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:55:54.682167Z digest=sha256:a3480fe4a5a945b4de85dd45a8fd7e37eff7b522ab0509cd2869b249b54aafb8

Observation 858b08f1-d929-47a5-bf18-0706f073017e · inbound

PRGB Benchmark: A Robust Placeholder-Assisted Algorithm for Benchmarking Retrieval-Augmented Generation cites this paper.

PRGB Benchmark: A Robust Placeholder-Assisted Algorithm for Benchmarking Retrieval-Augmented Generation Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:47:45.647885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:47:45.647885Z digest=sha256:79ad79e3a1b5e065f8619b4a51c5af01431a65ed1ad832b87a3d58497e40b464

Observation 9e88a9eb-a9f0-4166-b1c5-21f292897a40 · inbound

VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models cites this paper.

VLegal-Bench: Cognitively Grounded Benchmark for Vietnamese Legal Reasoning of Large Language Models Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:38:34.169422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T21:36:24.376401Z digest=sha256:0b26a9daeda65370444c448027e5392edbc819b19799a7ff090af06e741b85a6

Observation 45e1bd51-d266-44b2-bcc8-409a45e9d33f · inbound

Beyond the Parameters: A Technical Survey of Contextual Enrichment in Large Language Models: From In-Context Prompting to Causal Retrieval-Augmented Generation cites this paper.

Beyond the Parameters: A Technical Survey of Contextual Enrichment in Large Language Models: From In-Context Prompting to Causal Retrieval-Augmented Generation Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:03:12.194813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:02:02.300824Z digest=sha256:56e7d37e0273c38b5b63d9134d6fdc5b7dd5635429fa1b5dd347993b8d03f5a4

Observation 74d8f0b9-40fe-4346-9836-72879ab207eb · inbound

An Annotation Scheme and Classifier for Personal Facts in Dialogue cites this paper.

An Annotation Scheme and Classifier for Personal Facts in Dialogue Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:25.334681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:53:28.491889Z digest=sha256:de4fc770a4b6320a400575638b0a635d0d4d7cef88df942bb7c8f0dec4ab8da8

Observation c6d22939-05f9-4a1a-8605-c567d774fe9e · inbound

ProvenAI: Provenance-Native Traces of Evidence in Generated Answers cites this paper.

ProvenAI: Provenance-Native Traces of Evidence in Generated Answers Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:59:56.184397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T01:13:29.166367Z digest=sha256:a9d096e6c9631a08cc700c512d6775acd6e1ef0a1240c4d2c345bdd0a2c2b421

Observation 2f416189-c012-435d-aa98-57777ae7b434 · inbound

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents cites this paper.

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents Benchmarking Large Language Models in Retrieval-Augmented Generation

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:15:48.392474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T03:44:51.320606Z digest=sha256:32029def700e5634ba4d74b26b3bf9bee3d9614c94c97b5ec2d984866fe161c0