Pith. sign in

Paper Citation Record · LEDGER

KILT: a Benchmark for Knowledge Intensive Language Tasks

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2009.02252.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2009.02252 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:20:01.103232Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ae6e2ffe-b05a-4dfc-bfca-36565f6c260e · inbound

Atlas: Few-shot Learning with Retrieval Augmented Language Models cites this paper.

Atlas: Few-shot Learning with Retrieval Augmented Language Models KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:48:43.247472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T13:48:43.024120Z digest=sha256:9c5256768c844fa9e41cae78acb8c1a040aef232bb2b67fa05357b8c7547c3df

Observation 34153623-5d21-4be6-9d28-960cd80a6a80 · inbound

ART: Automatic multi-step reasoning and tool-use for large language models cites this paper.

ART: Automatic multi-step reasoning and tool-use for large language models KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:03:06.125700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T19:03:05.597295Z digest=sha256:cb7c77a9fd3ad9c3044b7dd47a1ed078f94dd2fda178aa5cfefc3d83e36b6ca8

Observation c985f481-cea8-4282-a398-961788db2fca · inbound

MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation cites this paper.

MRAMG-Bench: A Comprehensive Benchmark for Advancing Multimodal Retrieval-Augmented Multimodal Generation KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T23:20:01.103232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:20:01.103232Z digest=sha256:b5a7f500b30bdd10c6170f46de17fb35420fefe38453a9d7b84b38a6b7fc8649

Observation 6908e972-4569-4c8a-851f-9ec6e21f8650 · inbound

BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese cites this paper.

BrowseComp-ZH: Benchmarking Web Browsing Ability of Large Language Models in Chinese KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:04:49.993445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T22:04:49.915916Z digest=sha256:c37ecfe2062d81133a231a337a57dceecc7ccb3b2fd370213ebba8b83d85ef9c

Observation f00a0c21-6089-43e1-99b2-19f84d67a7e7 · inbound

ComposeRAG: A Modular and Composable RAG for Corpus-Grounded Multi-Hop Question Answering cites this paper.

ComposeRAG: A Modular and Composable RAG for Corpus-Grounded Multi-Hop Question Answering KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:51.315087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:51.315087Z digest=sha256:9382f656cd5d1655a1a37eccfbee0d57e7e46ce0935922cfee296a9105b3e2ba

Observation 7a926705-bc2a-4de6-87bd-b4a1fa1fe562 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 134

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.368308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:7cac5c42daaa2271de3040b7d156c089fe6813cc499f830178a87b27823a8928

Observation fd4fa678-d434-4089-b8b7-f6e865bcd201 · inbound

FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation cites this paper.

FlexRAG: A Flexible and Comprehensive Framework for Retrieval-Augmented Generation KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:21.182581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:51:21.182581Z digest=sha256:d2f78da521b0fd9045604b87d337d75865b05fb3837bcb65dadfabfdf677816b

Observation a281a326-5330-4eee-8dcc-f4ef1fdd1fc2 · inbound

MobileRAG: A Fast, Memory-Efficient, and Energy-Efficient Method for On-Device RAG cites this paper.

MobileRAG: A Fast, Memory-Efficient, and Energy-Efficient Method for On-Device RAG KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.450724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.450724Z digest=sha256:95468df687ba234ae1888bcb92f5985b898c8b279f619a86884bb5e57e29f7d3

Observation 783135f7-dc37-4956-9c78-47fdadcc6f27 · inbound

CROP: Circuit Retrieval and Optimization with Parameter Guidance using LLMs cites this paper.

CROP: Circuit Retrieval and Optimization with Parameter Guidance using LLMs KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:42:00.477274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:42:00.477274Z digest=sha256:601fc306fc9462aff54a00ce58e203f4b38f2abbe317b14dd08c07b37e44aead

Observation c7a1bc6e-b262-4a80-9290-ff64cf0b0dda · inbound

DiffLoRA: Differential Low-Rank Adapters for Large Language Models cites this paper.

DiffLoRA: Differential Low-Rank Adapters for Large Language Models KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:38:38.065081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:38:38.065081Z digest=sha256:ecc78c30cb31a780e8aa198261d50a1a2216df879c9d87a1921e09fa07659af6

Observation 47899f1d-20f1-41e7-843c-abc9cb0784e1 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:17.568390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:17.568390Z digest=sha256:26b4777f63d50fdb40a34d2920348886fcc2e76d752768aed390e9e1f92de017

Observation 8b4833a4-9b6e-4a28-b69e-692e9f712cfc · inbound

HF-RAG: Hierarchical Fusion-based RAG with Multiple Sources and Rankers cites this paper.

HF-RAG: Hierarchical Fusion-based RAG with Multiple Sources and Rankers KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T11:25:04.866356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:25:04.866356Z digest=sha256:946b58d540a91eea7f1134d4f888248462eacdc003955095fec361c50b9ceb4b

Observation 5105116b-8f9a-40e4-b370-78d7fa947479 · inbound

ARC-Encoder: learning compressed text representations for large language models cites this paper.

ARC-Encoder: learning compressed text representations for large language models KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T08:28:51.584949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:28:51.584949Z digest=sha256:a12daaeafdfc7673f8fa89b44d14661ea448ba0f32561a018ed44077592a3da9

Observation 8517c9b5-6ad2-4739-a3f7-ccea2bf6f0fd · inbound

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics cites this paper.

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.050570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:07:39.648269Z digest=sha256:e2c3b3faa11432a9f63decbb74898a2bf157802ada9b6516e80abf288eeaa63d

Observation 1053d484-92b5-46a9-8fd3-c8443271ea6d · inbound

How Fine-Grained Should a RAG Benchmark Be? A Hierarchical Framework for Synthetic Question Generation cites this paper.

How Fine-Grained Should a RAG Benchmark Be? A Hierarchical Framework for Synthetic Question Generation KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-27T07:30:41.516362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T07:12:41.103643Z digest=sha256:89466fca94c65a2aca9511ce98c70af8f6969ad974a5f492bb196aaeb9da1164

Observation 5fde8292-8d99-40d8-91b5-830fa53be35e · inbound

Uncertainty-Aware Hybrid Retrieval for Long-Document RAG cites this paper.

Uncertainty-Aware Hybrid Retrieval for Long-Document RAG KILT: a Benchmark for Knowledge Intensive Language Tasks

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:38:28.932226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T07:01:03.093718Z digest=sha256:6c81c61444f9de6f3fa3afec5c76ef327b8b87457ce111b64e5c2a25dbd13a86