Pith. sign in

Paper Citation Record · LEDGER

MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2501.03468.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.03468 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:32:57.233051Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b814be3e-e3fc-48e9-8bed-c17ed713b0eb · inbound

Benchmarking Poisoning Attacks against Retrieval-Augmented Generation cites this paper.

Benchmarking Poisoning Attacks against Retrieval-Augmented Generation MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:57.233051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:32:57.233051Z digest=sha256:3632e8a59d2522b876e2b20129f11166daf8439f1f39ab81a567f15b85b28141

Observation 4ae1502d-ea26-472a-8500-deb85b77a4d2 · inbound

Conversational Search: From Fundamentals to Frontiers in the LLM Era cites this paper.

Conversational Search: From Fundamentals to Frontiers in the LLM Era MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:26:11.806009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:26:11.806009Z digest=sha256:9ca582b42ef432e718d2ba50d5fd5dfc3bf6050a202557f674203a680b8fcf90

Observation deaafe6d-f287-4051-94f7-1b6fd245c8f9 · inbound

FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation cites this paper.

FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:21.117679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:51:21.117679Z digest=sha256:8f8cb59757ed6bf65376a23da2256ee8b877b9dcf7af9dd5d37fd2bbe49a7dd3

Observation eaf28562-69aa-4065-8aa8-e415c99fe5d4 · inbound

Granite Embedding R2 Models cites this paper.

Granite Embedding R2 Models MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T15:54:20.962573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:54:20.962573Z digest=sha256:3271a4a903d099d6ad598fef4fbeb980034a356476f009c00bbd63c443e5e145

Observation 5238b50a-f371-49ba-ad7e-6cdc8021b25e · inbound

DQA: Diagnostic Question Answering for IT Support cites this paper.

DQA: Diagnostic Question Answering for IT Support MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:50:47.679550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T19:46:34.294601Z digest=sha256:9afc9b7110d3bce2ed9839710907db1d4f5e1c71cf05e9efb541164994014e0c

Observation 49afdafa-e9a1-4435-8070-5dbbea83fb5c · inbound

RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation cites this paper.

RAG-DIVE: A Dynamic Approach for Multi-Turn Dialogue Evaluation in Retrieval-Augmented Generation MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:17:40.171386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T09:13:55.204404Z digest=sha256:4f14c7c73c68ea06d25b5c795688f9df854eda8a64baeda15165c138692d1e09

Observation 5bf18efb-152a-450e-9b1f-93252dbc4f01 · inbound

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs cites this paper.

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:13:44.850960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T17:06:46.379906Z digest=sha256:2c840ee364189000e9799224869285f00b18550e23350bf04b5599e0fdc293c2

Observation 1093b867-60f0-4523-863f-37d7bb1f057c · inbound

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer? cites this paper.

Grounded Cache Routing for Retrieval-Augmented Generation: When Is It Safe to Reuse an Answer? MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:23:45.149387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:16:48.588593Z digest=sha256:0a72a429317439b64e7a8e45296bb098706a41b724a82f4b9cafc1dd554b20e4

Observation c3631ab0-ea34-4e86-abff-9c835529e89e · inbound

5ting at SemEval-2026 Task 8: Strong End-to-End Multi-Turn RAG via LLM-Based Reranking and Faithfulness Control cites this paper.

5ting at SemEval-2026 Task 8: Strong End-to-End Multi-Turn RAG via LLM-Based Reranking and Faithfulness Control MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:54:40.223836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T10:06:44.776800Z digest=sha256:2ecd0874e34a5ed923e8f6de4942f5462a0cf39a3ef392f5db76d1dc44ce4297

Observation a214c589-612b-4642-92c6-f65755d7ba85 · inbound

Little Brains, Big Feats: Exploring Compact Language Models cites this paper.

Little Brains, Big Feats: Exploring Compact Language Models MTRAG: A Multi-Turn Conversational Benchmark for Evaluating Retrieval-Augmented Generation Systems

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:19.330776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T06:18:11.792566Z digest=sha256:ca830193fb8d5666ea6dd53260e4a28cefa3f979c233a7ed41c22f8aab436519