Pith. sign in

Paper Citation Record · LEDGER

The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2308.16884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.16884 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:07.811099Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T06:25:27.841513Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 148813b0-329c-4d88-a404-310f1eb67be9 · inbound

Qwen2.5 Technical Report cites this paper.

Qwen2.5 Technical Report The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:25:27.844262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T06:25:00.376073Z digest=sha256:585badb5742829d619af414438c6d4b5abb9015e85e6a1895c99ab06dc22344e

Observation ae538013-3ffe-476a-807a-4989d90432be · inbound

Qwen3 Technical Report cites this paper.

Qwen3 Technical Report The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:28.449337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T06:35:27.813995Z digest=sha256:5e5df90256a98743306db17cbe4b795feff1d2af5ddffc5268341c467eeff5cd

Observation 5d41d539-f2b3-404b-92c1-b6a48ae8adda · inbound

Cross-Lingual Optimization for Language Transfer in Large Language Models cites this paper.

Cross-Lingual Optimization for Language Transfer in Large Language Models The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:07.811099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:07.811099Z digest=sha256:c1493b935f8c4fc9e03ddcc7265bc5ef4bdc4df3d0320bbafaea25593ee50263

Observation 1d8a31b4-d56b-4e9e-9abc-43af9b76335d · inbound

Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text cites this paper.

Lost in the Mix: Evaluating LLM Understanding of Code-Switched Text The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:09.509245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:09.509245Z digest=sha256:d68edb0f6524015982808531c9ea493529116d1cec08b7ec837b2ac9048af137

Observation 3d170110-4d0f-4283-a254-f16331871022 · inbound

A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension cites this paper.

A Vietnamese Dataset for Text Segmentation and Multiple Choices Reading Comprehension The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:24.093318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:24.093318Z digest=sha256:ef749bffbb41be56e1be322755ddc33f0c974c0d7709a9932eca6233a34b6591

Observation 6cdb0b51-9bfc-47a8-8ffb-fd49eeacacdd · inbound

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems cites this paper.

From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systems The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T05:32:05.794157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T05:30:33.121799Z digest=sha256:a43ecb9579235b861a54d202f7168fcdc2e763ab236c941537c2b3a1ba7f456b

Observation e2c362b1-a10c-423e-82e7-6427b600c90d · inbound

Improving Korean-English Cross-Lingual Retrieval: A Data-Centric Study of Language Composition and Model Merging cites this paper.

Improving Korean-English Cross-Lingual Retrieval: A Data-Centric Study of Language Composition and Model Merging The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:40:51.024145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-22T00:39:24.748381Z digest=sha256:0bef0e32e93d141494eb41998786a0454a3969e8dc644bb02a623cd2285ab27a

Observation 1b0a24ad-03f2-4b3a-995b-dd68e180a1cd · inbound

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications cites this paper.

AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:28.791709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:26:28.791709Z digest=sha256:0e61956f1683336d04501f890037e5b7b1bbc465672d05f3e953bc29723f3e90

Observation 2db290fc-bd00-4632-983f-50814d13c33d · inbound

A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$\Delta$ Integration into Upcycled MoE cites this paper.

A Data-Efficient Path to Multilingual LLMs: Language Expansion via Post-training PARAM$\Delta$ Integration into Upcycled MoE The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:13:13.604415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T11:09:22.027588Z digest=sha256:17a93b1ea6def73d450847586478fdd30b73e4ad01936d469685a9ff196e55f5

Observation f372923f-1309-4463-829a-c640d0853e9f · inbound

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations cites this paper.

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations The Belebele Benchmark: a Parallel Reading Comprehension Dataset in 122 Language Variants

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-31T11:25:39.532395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:25:39.532395Z digest=sha256:3280ffa416067211f6e6f1ee5ba45d5dfeddafd4dfd68e13f0d2750503ef0368