Pith. sign in

Paper Citation Record · LEDGER

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG

As of 18 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2607.10626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.10626 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T10:22:55.558755Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

24 of 24 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 56d8b4d3-18be-4c85-b3c8-ad6090cce019 · outbound

This paper cites FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:e2ec4ccfa7535cc33b4aca34dd385d203110f0d1a0e19c0e91ee13159f7e4db7

Observation 668510e8-86fd-4af3-b654-6ec1f6a95337 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 2

Resolution
verified exact
doi, observed 2026-07-14T10:30:24.951785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:427b1c45d011328ac8293aae3f92614c1554776c8e8a233c0f3279f77f7ea1a4

Observation 79842aee-f0ec-4b0f-9806-0a63cbde44e5 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:d458ae3c4be5838465117ba3677eb3041279c750bd111e4fee00b47f04725f52

Observation 062ba047-64ef-4fe8-a791-8ebe6c161713 · outbound

This paper cites Coman, Ionut-Teodor Sorodoc, Leonardo F.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Coman, Ionut-Teodor Sorodoc, Leonardo F

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:bc893d41f2954afd9c0e69875af0a9556399b09252a517512fafb99cd10ad82f

Observation 6245d437-22e5-4d1d-8110-74344e6c170e · outbound

This paper cites Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:3e200ee5fa03418b96c505216e22a779b8bf07b88b11681afd93f5bd3fbb571a

Observation a9390467-7df5-4809-b694-4317c9b2ae50 · outbound

This paper cites Ragas: Automated Evaluation of Retrieval Augmented Generation.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Ragas: Automated Evaluation of Retrieval Augmented Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:91ab2792bb43566162e90cd74cf69226dc9dae05adabb83a06c46e4ffaea6f51

Observation 476f8506-8794-4e05-a692-97b3dd17c9f4 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:78e583a667660fd9cc649dc113df57d3d7c1cff0f798065b0a8f70c111184812

Observation c9bf7bb1-45ed-450b-833c-918b1c384b45 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:f5e894abfa826909c9dc79dd59df4d2593c9a5a1d2718f93f2640f69314ec5d7

Observation 72d0cc1b-c5e4-45e9-96d8-d53ab0736cf5 · outbound

This paper cites Prometheus: Inducing Fine-grained Evaluation Capability in Language Models.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Prometheus: Inducing Fine-grained Evaluation Capability in Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:634606051b1803229d9eb23db10fc55e6544297c3c5e96fa4a177d3aba825f66

Observation de0313bb-4bc5-4feb-8898-d65e9457cb67 · outbound

This paper cites u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG u ttler, Mike Lewis, Wen-tau Yih, Tim Rockt \

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:3c96c81a60239590e1c25982421c2deb2d2fed6214432592573a3c7b3300e293

Observation a711624a-1c8f-4703-a639-d87a258427ed · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:7f73de097bf33f2e27a5a61a7d1de078077cb6b98c2d40480d2ebdef06792196

Observation 5570a3af-a771-46d3-adad-a1871e240281 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:5119f6a0b23369d0f7bc9703a33f7abc8cdbf91b8af36c93b024b9cf9ea0c98f

Observation 3c110c5d-90f1-45fd-bd2c-3022b314d510 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:30ce35f6541683e35ecdf9a791b4452aba5c09cfdf9aa7697ce378dcebe69c35

Observation 63b30c6a-cc2f-46c1-8764-a0da85c4e3cb · outbound

This paper cites RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:45ea08091828694c28de1a77555a994dd7e6483640d2db733dbf0861e8e22f97

Observation dd37bc81-e830-48b9-b5bd-0ec954aba5ac · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 15

Resolution
malformed identifier
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:707dde6e2c39d93ce19b9d070cb25b86de06305985bca50e0c36c6397d034fc5

Observation 4e1154b0-a21c-4cbe-8c7a-0f8ae5f88e8a · outbound

This paper cites ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:88c3e0d5d69899e5899a176359272f317bbd0f36acfb81d81eba7f158f63350d

Observation 2b989eeb-3ff0-4560-bf38-917b3338ec6a · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:58a3043ffe0293025c061b40b03e60ff459e847268318fc5b41e1be0eed1b69b

Observation e5717006-fb16-4ca2-a301-403c87071cba · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:cca30cbefb977afcf40652531aaf53c8a1e1e612dbf3dc56c0ff20926bb3d02e

Observation 2bfa219f-0831-4e2d-8252-5e3a3f8ba9d5 · outbound

This paper cites Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:db40640629232168a2b51c540914988e4e979ccd1ae2bfb440bc565b77d04b9d

Observation 77399758-2e3c-4b3a-bd63-8061acd33a9e · outbound

This paper cites Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:e32c01bd07c2f3e5862703f040693eaf66a1787cc0a714cd6196427d1ee6f594

Observation 396578c7-75a8-48e0-9f5e-ebf93919abda · outbound

This paper cites Self-Preference Bias in LLM-as-a-Judge.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Self-Preference Bias in LLM-as-a-Judge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:292561695a479ce12d45cb040125dc71b5b3e4a9e0e7facb6e4f052ed50c4ba5

Observation 2894a54e-0b1f-4050-9046-4945156666b2 · outbound

This paper cites an unresolved cited work.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:c1cafd5763faf5117c30296161d947dbeb607af1bd825c1c73360f6bccca7139

Observation ba6c6b92-bfdf-4b01-a20a-8d4249cd5bb2 · outbound

This paper cites Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:88c7b37b0bae64c7d69a4a50fcd54b3a7934e18c2df64d666e5340b8259ee3d7

Observation 8e9ccb93-808c-4dbb-969b-64e95579dfaf · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Eval-Pair Matrix: Answer-Paired Meta-Evaluation of LLM Judges for Grounded RAG Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T10:22:55.558755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:22:55.558755Z digest=sha256:4ac0dd5bed09a09a1325c344bff1e05a62942a76f29037e957b859571ee0ea00

Pith citing papers

No inbound Pith citation observations are available.