Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis

As of 15 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 2 inbound Pith citation observations for arXiv:2507.23248.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.23248 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:58:56.261866Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T20:27:41.561992Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 0daa7c16-7838-4695-817e-8cc5abe79183 · outbound

This paper cites online" 'onlinestring :=.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:53.395639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:53.395639Z digest=sha256:2c404565f9f0d3062398c7a1499692e4a2a6ce8fe597321b0aebd05ca4eb493a

Observation 68116c71-cfb7-4ae8-949d-fb66c0835a5c · outbound

This paper cites write newline.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:53.477745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:53.477745Z digest=sha256:b5a327b5906869217daf91b72365127cfc709f0f86dfbead02030a40ac497417

Observation 2a78da93-eb7e-4e4d-b3b1-9d1aa326891d · outbound

This paper cites Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Bridging the Bosphorus: Advancing Turkish Large Language Models through Strategies for Low-Resource Language Adaptation and Benchmarking

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:53.612462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:53.612462Z digest=sha256:f6dbe270ca3db3e1ccaf83f447b3629628f84459017e994f099e830eb8c0d250

Observation 858acecf-7d3f-4b9f-b98d-96d4eec9936c · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:59.095173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:53.699216Z digest=sha256:154cd1f2ccf9e316220aa1392de6921eea73e986271ee830c84401a729debcf0

Observation 6dbadae3-daf4-4ffa-a576-bd026e0cb086 · outbound

This paper cites BanglaBERT: Language Model Pretraining and Benchmarks for Low-Resource Language Understanding Evaluation in Bangla.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis BanglaBERT: Language Model Pretraining and Benchmarks for Low-Resource Language Understanding Evaluation in Bangla

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T10:58:57.045200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:53.843745Z digest=sha256:f5ea761ae8ed1642a9a9256b20e73a9e5060574d92795c6707137f4157b327ee

Observation a50b5484-1c96-43ec-8679-01c4fddc9f5f · outbound

This paper cites Getting the most out of your tokenizer for pre-training and domain adaptation.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Getting the most out of your tokenizer for pre-training and domain adaptation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:53.975249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:53.975249Z digest=sha256:2e3e59d63503c367eb9d8eb6e5aeae415f7624962dbf3e905bf7a3c6f1a82e8e

Observation 1e6dcb52-f189-44d0-910b-82a8c38d0a5c · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:58.801468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:54.125148Z digest=sha256:8d4a71f40ef3b0d5a6ae0704862ffcd574e84708a7491819c12a7c17986bf2eb

Observation 9e3cab0d-2e1f-479b-b809-0b292d7fe3e0 · outbound

This paper cites The Llama 3 Herd of Models.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.238082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.238082Z digest=sha256:987a5788817cba6cd34b7c1ec5be937f8a11061d9f39a1999a28f45ed9be00a0

Observation 051a6815-7453-4e10-af86-424c126388d2 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.358014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.358014Z digest=sha256:3b2c226a1ca5ad86dccbac9daca2844ec08416f3e0aca87b5892707d29369cad

Observation 0a35c3d4-3eb7-4f4e-9c09-c7c5061d6b66 · outbound

This paper cites Mistral 7B.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Mistral 7B

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.522649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.522649Z digest=sha256:6c7907f71d18911f4136dd0d7550bfe5984124bd1bbec1530fe40f155e718406

Observation ea8743cc-968e-4944-a1e9-5e4acae32fb6 · outbound

This paper cites BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis BenLLMEval: A Comprehensive Evaluation into the Potentials and Pitfalls of Large Language Models on Bengali NLP

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.604605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.604605Z digest=sha256:2fcaf2568571d57ab62fecc6a3c345df7b78a68bccb405cbb7f8a4d6dd874bb9

Observation 1c805f7d-2954-479d-81ed-3167da3d3f13 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.770693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.770693Z digest=sha256:c64ee6de00d8dba7a4734b157ac2af9992300d6c04e3c9290b484599121c9b8d

Observation 24f91660-4aa7-4539-8bba-82b835c335bb · outbound

This paper cites Common Corpus: The Largest Collection of Ethical Data for LLM Pre-Training.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Common Corpus: The Largest Collection of Ethical Data for LLM Pre-Training

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:54.941258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:54.941258Z digest=sha256:b0c4f05f970ebc69f9ee588a20989a21f9395a375c05c1b65b69c2fe766af391

Observation 8882124f-d31b-43b7-bdc0-20c375a50af8 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:58.529427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:55.114652Z digest=sha256:3d5caf2b32e7ee6b047226bb86a9eef3207fb9f124b54d851ff1bdee51e0eb71

Observation c45f8920-b22c-4859-a5f2-6b910cc3b664 · outbound

This paper cites TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:55.250378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:55.250378Z digest=sha256:0511f19f84cacdfcf45c963d8f114b1fa394ae39355515d3d1686fa8484c949a

Observation a7c581f4-c124-4866-b441-fb91b1ed3955 · outbound

This paper cites Qwen2.5 Technical Report.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Qwen2.5 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:55.309685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:55.309685Z digest=sha256:deed4503f01ace64f7d90b8e941f021351a1bfdb6490aa7b63b796335fd18b3c

Observation b6a453d8-f870-473c-ba4e-cc15a8bb616f · outbound

This paper cites TigerLLM - A Family of Bangla Large Language Models.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis TigerLLM - A Family of Bangla Large Language Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:58:56.605400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:55.473837Z digest=sha256:08dc978118f675dc5410bf2824560d9e0c1f2ba1e37324d97de78e45c6d6a9c3

Observation 744f229b-64d0-404a-97ad-1ffedb556a1e · outbound

This paper cites Shahidul Salim, Hasan Murad, Dola Das, and Faisal Ahmed.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Shahidul Salim, Hasan Murad, Dola Das, and Faisal Ahmed

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:58:58.148390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:55.589677Z digest=sha256:c7154466925f70b233b4d8473305ccaede419930f3f2f2b6a5154bf4f4f125dc

Observation 000c2b0f-d387-4c98-aaab-fb0f6aa18c03 · outbound

This paper cites BEnQA: A Question Answering and Reasoning Benchmark for Bengali and English.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis BEnQA: A Question Answering and Reasoning Benchmark for Bengali and English

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:55.698898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:55.698898Z digest=sha256:bafefbb9abd74e7b6a37177188e082b4aebcdb268e2760cb58249506fbf62238

Observation 9ec22d8d-f984-4e9c-bb0a-3610eb75b6b5 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:57.896082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:55.813844Z digest=sha256:1ebd78b4e2541b54da38d84491c0070778d3a071dc37ea9489f4e4f595deee5f

Observation 0a0d9a21-3ba8-4ceb-8357-784d7402d1f8 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:57.661497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:55.996306Z digest=sha256:76708e9ebba4b2a8fa2ebbb7dfcb38aab4d9381a0e089e56bb3ede8bd77968b3

Observation a254f121-31ea-4cfc-af80-eca4031dcf46 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:58:57.357810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T10:58:56.131277Z digest=sha256:ecb49a9899c8aee76ba16a0422c5909158cd8b2fb070f81da14f6341d054880e

Observation ac0cf0a8-70e3-4ef4-8b8c-381f2cac32bb · outbound

This paper cites Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model.

Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:58:56.261866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:58:56.261866Z digest=sha256:b1d28f71864977019d6bb16cc896e8d77ecccc5a664df1eef632d484e80db6be

Pith citing papers

Observation 628afa96-9cc1-41c4-a176-2eb96dfbf631 · inbound

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability cites this paper.

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T21:00:17.798234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T20:59:59.582440Z digest=sha256:9c104af932a56f903051b958fbd88053c8ddf2e6bafaf83ae7a2fe76420e2e9e

Observation 138739b6-0e0c-4dee-b519-406161a46a65 · inbound

Creating Multilingual Mental Health Dialogue Datasets: Limits of Persona-Based Localization via Nationality and Language cites this paper.

Creating Multilingual Mental Health Dialogue Datasets: Limits of Persona-Based Localization via Nationality and Language Evaluating LLMs' Multilingual Capabilities for Bengali: Benchmark Creation and Performance Analysis

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-26T20:29:57.597511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-26T20:27:41.561992Z digest=sha256:ceb0aebc65e7f2169eb24116f9893b29c4416a1fc51147d77b47cc4acea229b7