Pith. sign in

Paper Citation Record · LEDGER

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

As of 10 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 2 inbound Pith citation observations for arXiv:2507.08538.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08538 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:19:59.488074Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T22:16:42.836058Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:09:41.063430Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact2
  • verified fuzzy1
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ea5dc9a5-023e-4979-9cef-feceaabe33a4 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.116357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.286503Z digest=sha256:106da5b9b7fa0a71dfef249b31bee3f7dacb86efad8e367c1efb4562dfaa26d1

Observation a3f2fbe0-9388-49c3-8b8f-62c6a6591335 · outbound

This paper cites SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks SIB-200: A Simple, Inclusive, and Big Evaluation Dataset for Topic Classification in 200+ Languages and Dialects

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.297703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.297703Z digest=sha256:fd87aaa3151fd9cbacf5f5ee9d27bdd2c969b3a9d2129612e92edaf94f2351c1

Observation 74157662-1675-4467-8bcc-c0cddf292b77 · outbound

This paper cites IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks IrokoBench: A New Benchmark for African Languages in the Age of Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.301933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.301933Z digest=sha256:b0a0671599e126abd551c5ced986d033ac90191c79806237b08c329b1d5db49a

Observation c5d39465-3f03-4cb6-9ff7-818308f676a7 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.313039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.313039Z digest=sha256:274fdeb80c0db734119a2c7757f3df897d206b147278f63ef5fc418e034fa239

Observation 73d888ff-7b96-4575-bfc2-f15485c3c338 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.316902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.316902Z digest=sha256:41cc514437f7b5d427431b8a13afac58e1341ea4d5366ee51ab7214addf2d704

Observation 12900b22-1ea5-432d-9e9a-7d5c63184090 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.103872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.320491Z digest=sha256:b4ee92cea036176ed51d4dfb6bdbe843ae16d9fb7a0e01f59728fff50b735f4f

Observation f50d2a13-f26b-4a51-ab1f-a6eab9fb5177 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.324622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.324622Z digest=sha256:182c45e6e2831f0f172db44eb5e9ab1a7d4ce85f8f1e93da10718e1638132cd2

Observation 9495b29a-229b-43fe-b043-f51a8d8d39c5 · outbound

This paper cites Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Uhura: A Benchmark for Evaluating Scientific Question Answering and Truthfulness in Low-Resource African Languages

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.327369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.327369Z digest=sha256:26bc6189f12eef67069abea77a94b63202a57395f3d4f26dd11660f47bc723b9

Observation b13a21ec-78e8-4094-977a-93b2acdc80f9 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.331114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.331114Z digest=sha256:71a45b51d8866369f4903fae0fec9e79e10af75ceb47351ff19107e40ffb3ceb

Observation 316ba190-ca4d-4be0-804d-3cb9a2a4f783 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.345260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.345260Z digest=sha256:175a63f949254fe547b4ea1bca2b57ba3a17913485b4317d882339ed08eadf4d

Observation 2dadaf57-a4d2-4e71-9310-6510fe1dfcf5 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.348657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.348657Z digest=sha256:eac3bfec211a4842b571036a311a634f1550e57c64170073e421010042631b8d

Observation 2ccd1855-5476-4a80-a4e0-891fa333c7cc · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Training Verifiers to Solve Math Word Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.352217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.352217Z digest=sha256:cb565512c7c2220b5bd3e544a60d3c403780f8eccd6f8fd903d8f0babc1f178c

Observation 40d3f81d-f183-46b7-9f57-013f7a0f15e9 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.358937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.358937Z digest=sha256:9177087197fef03df6c12e4a34fdb258fa28194b2579da5cdd6b502a2730f39a

Observation 5b815c6d-6f48-449e-a8d6-69bbf1ee79f4 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.362269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.362269Z digest=sha256:4272658fc9095aae8bc913e895dbf0d4557452662a5d0e90df82662adcaf9c8b

Observation 68de6cc5-cffb-4362-b86b-306e8654b6e7 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.370724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.370724Z digest=sha256:4cf0a25f5f1a0d377971dae192e34081c62984843188b751852ffc9d6db1496b

Observation a3ebffed-3383-44b8-b7c6-1c8850839acb · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.373566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.373566Z digest=sha256:8a810eb7c7b35d14eaee0275c80607d179d29db7005231c76d451e7abfe8f405

Observation 40f01db5-8738-4e51-a15a-31e2aabac34a · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.376093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.376093Z digest=sha256:df8c582bfbc74100dda828480556c1b5a45a61b8f6e33476deb8f493a6e99247

Observation 2d44566c-ee72-4965-bd9b-62fbb6cefabf · outbound

This paper cites Eberhard, Gary F.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Eberhard, Gary F

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:20:00.084843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.380102Z digest=sha256:6fbd930e8af0fbe443e24d7c0ff494418475b3499603f4cbe3e22deb1c5ee904

Observation 371f5463-c200-41bf-a438-ba27692daa18 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.065614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.387423Z digest=sha256:06280495cbb429c6c0969def9dc156373db4ee3e17c777f44a9de8ed60957a5d

Observation 7fa8c939-2e20-4dbd-a464-da82de0f37d9 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 21

Resolution
verified exact
doi, observed 2026-08-06T18:19:59.597764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.397423Z digest=sha256:d9abeaf76a9711b887b07f34ab80a7e5cc9281fcd8f61ca1dd68485712e28db2

Observation 1d22288d-4ec5-4ebf-8937-6ec2001dda73 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.401528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.401528Z digest=sha256:dadc0f0126bbf058be574564b6200819219cd94defd1f161e7c97ab662b50467

Observation d646cb52-a33e-4b21-9279-fee469173916 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.404568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.404568Z digest=sha256:5ab250298766a69f3a693b2eff8aba8628bf5133b86dc09ce68caaa5aa42ef25

Observation 2abf76eb-59de-4ed3-b413-675fdb541f32 · outbound

This paper cites The Llama 3 Herd of Models.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks The Llama 3 Herd of Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.420042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.420042Z digest=sha256:2637bd9cb7b88de249e4aceb03b424ad8441e6848b037864807f0e904e375d7f

Observation 2a3720dd-1d5f-4fa7-a27b-a521650b0d69 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-06T18:19:59.572001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.422684Z digest=sha256:628add2ebdeb87bed31def1c3743ad92a533e4c75e8937e54fafef5eb326871f

Observation 3f6dd766-7819-4598-9c5c-3e5fee4c167e · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Measuring Massive Multitask Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.425715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.425715Z digest=sha256:deeac0bdbd6c571ae3bf7aa5e220e503db8ab3069a5756477bf49c17796dbb82

Observation b0ccd95b-80db-4b7d-ae55-6aa4eaa3fd47 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Measuring Massive Multitask Language Understanding

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.429610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.429610Z digest=sha256:119cb2f3aa9045dc265e50e4200731817f10b815d87fd09812313bc9f643e338

Observation 9f4be4c7-7d94-47bf-b8bf-a289c9cc45c2 · outbound

This paper cites Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.432414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.432414Z digest=sha256:0175ba08f4cd000d334822fd92402f89545cc77102fb32801bab5f1aaec67f8b

Observation 6abaf680-c7f8-44c5-9d50-82fecfaff93b · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.435453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.435453Z digest=sha256:c50bb6f351eb7ed5636d30cc1d7338916e20d9509637112aab06dd0926e1d628

Observation edc97634-f5c4-44aa-b52d-5005161a89a1 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.049239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.438133Z digest=sha256:09b97b9d73f9234d764bf25aae43950265cb97de4f111e70180e0c2168ed292e

Observation f2fa17cc-eb8c-4d46-b7b0-47104a8c9d1c · outbound

This paper cites AfroBench: How Good are Large Language Models on African Languages?.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks AfroBench: How Good are Large Language Models on African Languages?

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.443625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.443625Z digest=sha256:42aae6cdd9bd5cc45164d3e4f61b33a20dbafec75625045d014de6c3c8f4879d

Observation 0d0cf63a-55ed-4e2c-a021-74a74d668996 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.446447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.446447Z digest=sha256:b08fd16ea55b6e9ce3a44be21533014f5b152ea09f686a3d1bc12b39821742a0

Observation 09b840e1-b651-4402-bccc-354a17892cfa · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.031481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.449243Z digest=sha256:7fcad31dd5f6a2edde317fe35d2498970891029bcb00e3c85454db406fdb44da

Observation ea9e2e71-2600-448a-b84c-7cbb86697667 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.018207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.451844Z digest=sha256:b69c9a23b40566db90ff75d2d6b4df9e5e3342e6d245813e953b9e94250dbdf1

Observation ec21d80d-d331-42aa-8f16-d7cbebd7a210 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.454603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.454603Z digest=sha256:84ddcf4d853384e76ab2cb7cb7159c74b4f724953b40e490d6d161a556ec2fb0

Observation 2fa77bbf-7f2e-4af4-9480-bbf1bc74644e · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:20:00.000793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-06T18:19:59.457204Z digest=sha256:59da711ff48daa4c73f2f348c8e1e3d40f311097ff25db0d412f1a4c45dc6aa4

Observation f4dbf787-c595-4360-92b2-1435ec378768 · outbound

This paper cites Language Models are Multilingual Chain-of-Thought Reasoners.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Language Models are Multilingual Chain-of-Thought Reasoners

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.459779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.459779Z digest=sha256:15f8fb4974c6f563456a545a058b3796b489103a411f82009b40548e967ad3b8

Observation 75c7a333-5511-4883-9966-214c158b7954 · outbound

This paper cites Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.462845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.462845Z digest=sha256:960532825b0fa79d62c76b93bc56cc15e6595e88629c571f4a386ce212e60a45

Observation c39204d1-abe3-4fc3-a480-15c1b4b0e12c · outbound

This paper cites Gemma 3 Technical Report.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Gemma 3 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.466179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.466179Z digest=sha256:7144e9a6b053d0f00a0c30b00189c7bf5e1e4bb6499fc1fbc4a5f5764fac0b83

Observation 9e9bfe70-e9d4-449f-82b4-ed9ee62f7b75 · outbound

This paper cites Towards Multilingual LLM Evaluation for European Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Towards Multilingual LLM Evaluation for European Languages

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.469267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.469267Z digest=sha256:448d557f6dcfdb6f0281b2f250ecdd4f71eb671cbc3e86119222d7b24971ce2e

Observation b138c725-3ee2-4ae8-94b3-f9fa36f6f61f · outbound

This paper cites Towards Multilingual LLM Evaluation for European Languages.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Towards Multilingual LLM Evaluation for European Languages

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.472849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.472849Z digest=sha256:aaa59fb63605c1fce751958694d5f69e3b478f10bac8767b399824378ac8304a

Observation 492a34b0-5344-4c66-b1ad-0411e78bc728 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.475927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.475927Z digest=sha256:5eff84a9cca75a39276ac87aa8eefa3c5a404ae3e30e889c584d25594e56facc

Observation 47f9d69e-db9d-46d3-8415-ca814bdfb838 · outbound

This paper cites an unresolved cited work.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.478444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.478444Z digest=sha256:d7119f461c57f7afa2ad706289a72ac4b3d6e533f9c8179da48fe7815d45fa31

Observation 92444df7-09a0-46c4-bbfb-5e41e53ca32a · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks BERTScore: Evaluating Text Generation with BERT

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.481633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.481633Z digest=sha256:6f677e0ea8ca72a8a2978b7160b0fe5e11f229f9f467cd2ca032c7682dc30104

Observation a693c908-41b3-4107-baa7-f147dc88334d · outbound

This paper cites online" 'onlinestring :=.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks online" 'onlinestring :=

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.484979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.484979Z digest=sha256:2e329bf588b8dd046ff9533a51bafd4570c2a49306979f536917e666022b9c52

Observation 15b845e3-1c0e-4bc5-a390-2e997e86499a · outbound

This paper cites write newline.

The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks write newline

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:19:59.488074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:19:59.488074Z digest=sha256:75056db7523916a7f47f6ca63c1247514ca1e63f553ae3493600dfbd08a54fcf

Pith citing papers

Observation d6b8e740-e173-4c1f-be05-b7e98a48b732 · inbound

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights cites this paper.

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:57:09.878647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T22:16:42.836058Z digest=sha256:0a60d2bc32486a46d37356a76c5fe4f074efda122cec9f0881ab713253062bf5

Observation fe845ad2-462c-4b79-9366-31d37fc2fe62 · inbound

The Language-Energy Divide: Measuring Energy Costs of Multilingual LLM Inference cites this paper.

The Language-Energy Divide: Measuring Energy Costs of Multilingual LLM Inference The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:41.065578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T12:12:45.853559Z digest=sha256:ba82dc65afe7e1f38c5c66ad93bdac76966b38ea53d1d514ee3bd06a354c6189