Pith. sign in

Paper Citation Record · LEDGER

BALSAM: A Platform for Benchmarking Arabic Large Language Models

As of 19 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.22603.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22603 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:36:00.280571Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:03:03.350371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:30:52.664737Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01da45ee-27e7-4952-83a4-46425ac1fb33 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.480710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.259136Z digest=sha256:2661802ecaad5ab32b9784d220b16443c93125a2f18851d9cb0dc4d215e0c147

Observation e06ab83d-1c55-4735-8fea-d57a18341fac · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.463511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.264374Z digest=sha256:3e5b82eb58235bdf08f4ecbcc55897bc599194ee475f0b9cb7260539254edeea

Observation ac4e2a38-4a84-45c0-83f9-18e31d69a0b2 · outbound

This paper cites Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.539955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.228330Z digest=sha256:fdde4576c81c69a7c3ba46c81d9b8a1fcfcba73bc6a489c9d005862bd7d64f81

Observation 157a8e67-c6f6-40ac-8138-da5596c037cc · outbound

This paper cites Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.239773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.239773Z digest=sha256:077be6ed4a402d161c28a40fd18ee19605a848c50f4ce48ed80448d3c303dfe7

Observation 1afaefda-ccae-4945-b94e-a2e33ff066bd · outbound

This paper cites Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.445382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.269899Z digest=sha256:bf91f3185493b99bfa549ff919ea720ced2e42cca41c60735cab382b2451457f

Observation 2e23c6ce-d51d-4c64-8042-f66350784495 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.427852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.275656Z digest=sha256:062dde2b2e2409ef62cda04f8d9d40416dc719d15c5a4069e101e33ea488404f

Observation a1be0072-2efd-47be-8c2d-b37fd6a6c9bc · outbound

This paper cites score": 3,.

BALSAM: A Platform for Benchmarking Arabic Large Language Models score": 3,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.410565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.280571Z digest=sha256:d35c8ce5aa9a737d8169953d5762795b9a2e0c52223237581d63c5fec46c2c56

Observation 32a637dc-4852-47fd-a38b-a61e37a53dee · outbound

This paper cites In International Conference on Learning Representations.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In International Conference on Learning Representations

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.521004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.233976Z digest=sha256:e17123d4dbb33cb8af0e578fe3fe7e7283ea60d80d4f0bd1b9cefd852846c75c

Observation 9c55c0a3-d89c-49d2-82b3-628f3da1c94d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.215303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.215303Z digest=sha256:f06cba479a7b44d23dae5cee196e0c3547ce7bbc96b7c42741498a4d135dcddf

Observation bd0921df-52e8-440c-92f8-00ef16da9de4 · outbound

This paper cites Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.250729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.250729Z digest=sha256:b77fa3b6cb6816eeedc584b9504114c72361b87adb6b22c476b238fb3cc5355f

Observation d3f97aaa-4e76-433a-8eda-725ae6fffd57 · outbound

This paper cites Creating Arabic LLM Prompts at Scale.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Creating Arabic LLM Prompts at Scale

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:36:00.366212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.222651Z digest=sha256:375917473db06055d41ae990af9ccbd329863eccedfe32163efb86987f2edc25

Observation 06ffc0bd-6a50-4701-a2e3-efbf85ba8cfe · outbound

This paper cites In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.500643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T11:36:00.245539Z digest=sha256:2f43b2a663b5df8680e8c3325a3ba12de854f6abb670fa1cad4b63f56c6ccf65

Pith citing papers

Observation d292673a-38ce-4876-b9e2-315e12e1b824 · inbound

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection cites this paper.

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection BALSAM: A Platform for Benchmarking Arabic Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:52.671051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T19:03:03.350371Z digest=sha256:3c79ce0f7b0cce44c20e243721ee8ce01f03a94763a011b587c874f231c75837