Pith. sign in

Paper Citation Record · LEDGER

M-RewardBench: Evaluating Reward Models in Multilingual Settings

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2410.15522.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.15522 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:09:01.006991Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T11:22:16.624253Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3f7a2503-723d-48c2-989a-e7ec16c32645 · inbound

JuStRank: Benchmarking LLM Judges for System Ranking cites this paper.

JuStRank: Benchmarking LLM Judges for System Ranking M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T16:59:49.490088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:59:49.490088Z digest=sha256:2098f3b814a652f3e94197d4de780c9042e933483382d6da9fb557367869c85b

Observation b82689fd-1e95-4763-a268-884fda411c31 · inbound

RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment cites this paper.

RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-11T12:54:54.513004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:54:54.513004Z digest=sha256:6f5a080b14ea497576583dff7c080bbbc19594db651f8dfd710f38699c49de6d

Observation 5daeb795-2c92-485c-bbf4-c56d8c29c6b2 · inbound

LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies cites this paper.

LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T11:43:59.596163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:43:59.596163Z digest=sha256:b71c093d0cc5296c00d64623442e9122e12133fc193e02d4e027c9ca7d98fdce

Observation 0af39565-f021-4e93-8634-6d84e6b00ba6 · inbound

InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model cites this paper.

InternLM-XComposer2.5-Reward: A Simple Yet Effective Multi-Modal Reward Model M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T17:18:40.289337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:18:40.289337Z digest=sha256:d7d157e0749c06f743a9884919c289c8c237b65a6a74daf6441329bf5e00476c

Observation b42ebc1d-4433-42cb-aa60-b979887cf0d4 · inbound

A Systematic Analysis of Base Model Choice for Reward Modeling cites this paper.

A Systematic Analysis of Base Model Choice for Reward Modeling M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:09:01.006991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:09:01.006991Z digest=sha256:f1dc8c5427c8e51b1762f3d84fed31c52aab3398176f1303702c5c3b2f57f5f4

Observation 95536eae-6b54-4d11-bdc8-9f8ea417005d · inbound

Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation cites this paper.

Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:42.658486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:22:42.658486Z digest=sha256:fdcbcf70ab51a02a69f0afa6cf49cecafe57eaba01fa52cfd620a97a2eb83081

Observation 17804e35-3018-44c5-ac0e-db440a1576e1 · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:22:16.627485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:ce12478d6905cde72040c7ba5224ed6db016872100e80dc6b7c5fd2f31534e33

Observation 96fd7ad5-8105-406f-b63c-6d9872513cd8 · inbound

One Tokenizer To Rule Them All: Emergent Language Plasticity via Multilingual Tokenizers cites this paper.

One Tokenizer To Rule Them All: Emergent Language Plasticity via Multilingual Tokenizers M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T04:25:12.734104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:25:12.734104Z digest=sha256:48a5fb192a1ff49d14b117cfaa3189dc210d4a4f025d8543d9373c3e54da4841

Observation 0ca32232-0d4a-4190-bdd2-eca796c144e0 · inbound

Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples cites this paper.

Relic: Enhancing Reward Model Generalization for Low-Resource Indic Languages with Few-Shot Examples M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T19:32:11.789138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:32:11.789138Z digest=sha256:1dfe5e4870cc62bba7134f6c142a4003b5ec33a568f3d868db0dd2ed8d68b88e

Observation 7b27f7cc-25e1-485c-9580-29ab4868c157 · inbound

Activation Reward Models for Few-Shot Model Alignment cites this paper.

Activation Reward Models for Few-Shot Model Alignment M-RewardBench: Evaluating Reward Models in Multilingual Settings

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:02:42.374334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:02:42.374334Z digest=sha256:f846e4a1caabb10c6ad35d12c5e437fc17e6882c48e20d9ba84d9f366ecf9327