Pith. sign in

Paper Citation Record · LEDGER

BALSAM: A Platform for Benchmarking Arabic Large Language Models

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.22603.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22603 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:36:00.280571Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T19:03:03.350371Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T23:30:52.664737Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01da45ee-27e7-4952-83a4-46425ac1fb33 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.480710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.259136Z digest=sha256:07dd32093b5720c34182f01070cd97b5d1098c11a113043c810558e63cd354d4

Observation e06ab83d-1c55-4735-8fea-d57a18341fac · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.463511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.264374Z digest=sha256:4c0f458624a1fc725f8fd30bc338d84735f937e3bc4ab42148e97ae88a8b5009

Observation ac4e2a38-4a84-45c0-83f9-18e31d69a0b2 · outbound

This paper cites Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.539955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.228330Z digest=sha256:eb317b6df58440e1f6e863af4e1e5145fc3726a4505a6635c99c3432a75e3e59

Observation 157a8e67-c6f6-40ac-8138-da5596c037cc · outbound

This paper cites Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Arena Learning: Build Data Flywheel for LLMs Post-training via Simulated Chatbot Arena

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.239773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.239773Z digest=sha256:f1e691c6174f018886eb2fe616f3a9eefb1658ad7f20db771ef91ffb8f3dbcde

Observation 1afaefda-ccae-4945-b94e-a2e33ff066bd · outbound

This paper cites Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Evaluate the generated output by comparing it to the ground truth, considering how well it addresses the original prompt

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.445382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.269899Z digest=sha256:0f260a8c18d5fc303677fb15efb4cc0cc76a29acf07e4b865de55e9d4cc92086

Observation 2e23c6ce-d51d-4c64-8042-f66350784495 · outbound

This paper cites an unresolved cited work.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:36:00.427852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.275656Z digest=sha256:bbd392a75e1b2467c278ca4b230967b176fd44c730b388e165a1eae23d0ac790

Observation a1be0072-2efd-47be-8c2d-b37fd6a6c9bc · outbound

This paper cites score": 3,.

BALSAM: A Platform for Benchmarking Arabic Large Language Models score": 3,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.410565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.280571Z digest=sha256:f156f7478da0fd8244ba219cf6e4874143099d12ea3ef08ccbb19029cbb2d4f8

Observation 32a637dc-4852-47fd-a38b-a61e37a53dee · outbound

This paper cites In International Conference on Learning Representations.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In International Conference on Learning Representations

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.521004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.233976Z digest=sha256:b169c7f96a2484aa852aa1396cf56960f05cc11a5f245716a41a6e4ead0be118

Observation 9c55c0a3-d89c-49d2-82b3-628f3da1c94d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.215303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.215303Z digest=sha256:f06cba479a7b44d23dae5cee196e0c3547ce7bbc96b7c42741498a4d135dcddf

Observation bd0921df-52e8-440c-92f8-00ef16da9de4 · outbound

This paper cites Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T11:36:00.250729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:36:00.250729Z digest=sha256:4195cd51efc7e80a9d207669fe878d76522b2cbb7e7972d2cc3b0ef7285bd340

Observation d3f97aaa-4e76-433a-8eda-725ae6fffd57 · outbound

This paper cites Creating Arabic LLM Prompts at Scale.

BALSAM: A Platform for Benchmarking Arabic Large Language Models Creating Arabic LLM Prompts at Scale

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:36:00.366212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.222651Z digest=sha256:c227013b530bb023db6423b515c193f6706b7455377dc4a24c54cd11fb790b6a

Observation 06ffc0bd-6a50-4701-a2e3-efbf85ba8cfe · outbound

This paper cites In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE.

BALSAM: A Platform for Benchmarking Arabic Large Language Models In Proceedings of the 31st International Conference on Computational Lin- guistics, pages 4186–4218, Abu Dhabi, UAE

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:36:00.500643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T11:36:00.245539Z digest=sha256:5cb5d32e9ce53d48e64944bc9568630e871b6099ae4a167109651d2ad8e6aa89

Pith citing papers

Observation d292673a-38ce-4876-b9e2-315e12e1b824 · inbound

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection cites this paper.

Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selection BALSAM: A Platform for Benchmarking Arabic Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:52.671051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:03:03.350371Z digest=sha256:79270ca25b94906df6db96c290ac40ccb37d5be52e2f081d1a22f3831db98d05