Pith. sign in

Paper Citation Record · LEDGER

Crosslingual Reasoning through Test-Time Scaling

As of 17 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 8 inbound Pith citation observations for arXiv:2505.05408.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05408 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:08:00.731058Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:27.736254Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved44
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 5c226e72-5596-4ff4-bac7-d5ce44e0736d · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Crosslingual Reasoning through Test-Time Scaling Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.347215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.347215Z digest=sha256:320d812e42e00851cdb01764753e22fbf1203645c2652f2f1cfa92fd79a44358

Observation d0c5a047-cf59-46e1-8efb-b890d9fe1e46 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Crosslingual Reasoning through Test-Time Scaling Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.354200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.354200Z digest=sha256:ae23bb30871499185b2abd267b7fcbbe36868f2361a8b1c13143d6be1aa99dee

Observation 516bf54f-237c-4363-9fcf-cbab9a04ec3e · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

Crosslingual Reasoning through Test-Time Scaling Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.359981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.359981Z digest=sha256:48b3ff14c51be9d9360ba2a77855926bdfcfa650d03c133a1a9e4153d739db19

Observation 94a9e3c3-f245-4b61-b516-ed6affaaa682 · outbound

This paper cites A Simple Model of Inference Scaling Laws.

Crosslingual Reasoning through Test-Time Scaling A Simple Model of Inference Scaling Laws

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.365205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.365205Z digest=sha256:29dfe929f6efb55c4f3c191804b5c881a6d2f8ae3be3becacaf9d249ba4bf7bb

Observation 7f1c115f-e392-4520-926a-75e357c6d802 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Crosslingual Reasoning through Test-Time Scaling DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.370482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.370482Z digest=sha256:ddb7c55fb958547a52cb8f383aac9929a0d4ed84d5ba87b7baad3842802f5a1e

Observation 1d84dd28-f3cc-417f-ae0f-79e4937e7343 · outbound

This paper cites OpenAI o1 System Card.

Crosslingual Reasoning through Test-Time Scaling OpenAI o1 System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.375882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.375882Z digest=sha256:d3321f017e5a4b9e7327357ebde6a05f3d48abdf0346747ba93206f483a5a9d3

Observation e371c722-799f-469d-854c-0011ac9a5235 · outbound

This paper cites Openai o3 and o4-mini system card.

Crosslingual Reasoning through Test-Time Scaling Openai o3 and o4-mini system card

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:02.029517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.382184Z digest=sha256:debc1274854164ced13b3349ebfe0fd498eceefc40a323acabfba25d200e655c

Observation 64d3d8b5-6666-40f0-b16b-fc6a942e6655 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

Crosslingual Reasoning through Test-Time Scaling Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.387298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.387298Z digest=sha256:3e15e52c50d435bb64e6cdfafc4dd950a134ab2da50f9a03966c03c7f35cfa3d

Observation d2d78fea-1f00-443c-9afc-7fbbba27738f · outbound

This paper cites s1: Simple test-time scaling.

Crosslingual Reasoning through Test-Time Scaling s1: Simple test-time scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.393139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.393139Z digest=sha256:2d7a61f56cf79d22692f4fe39f430397231728877a14e4cc762ed34e86b6f11d

Observation 1977189d-2ae8-47b2-aa8a-813a084efe6d · outbound

This paper cites LIMO: Less is More for Reasoning.

Crosslingual Reasoning through Test-Time Scaling LIMO: Less is More for Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.398411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.398411Z digest=sha256:f16af1bd4f1831d57fd5ddf8346dc6645f5d21feecac352a02d350963bcf4305

Observation 06ec4f4b-f10b-4e6e-9316-b1c7927185fa · outbound

This paper cites The multilingual mind : A survey of multilingual reasoning in language models, 2025.

Crosslingual Reasoning through Test-Time Scaling The multilingual mind : A survey of multilingual reasoning in language models, 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:02.008119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.403327Z digest=sha256:ece9a51027609d78888996dcf8731ea7d06c3ae8eb738d76975e2e3377846319

Observation 0acc2b47-6932-4b8d-9247-59192bbe8256 · outbound

This paper cites T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling.

Crosslingual Reasoning through Test-Time Scaling T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.407935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.407935Z digest=sha256:2b70be480953fc2314aee5f72d7be03987682dab32e479ca130c5bec613ef5ca

Observation eb9f13d0-58e3-4a2b-b398-ad4e7f3ebe08 · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

Crosslingual Reasoning through Test-Time Scaling Training Large Language Models to Reason in a Continuous Latent Space

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.413953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.413953Z digest=sha256:c326b0eb59a1e7f9c6dc84463dc124006708e04f75610f4c59b30c8ca5b03ba1

Observation cd4af6b3-da47-4e37-9b20-3cff53ccabb8 · outbound

This paper cites Critic: Large language models can self-correct with tool-interactive critiquing.

Crosslingual Reasoning through Test-Time Scaling Critic: Large language models can self-correct with tool-interactive critiquing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.991358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.419580Z digest=sha256:bd794b058bb996a3656adec8aeb09fc659b98d27f5b540b52307f219a0d38ce9

Observation 007cd9e1-c964-4d72-b0c1-324c490d9653 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Crosslingual Reasoning through Test-Time Scaling Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.424521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.424521Z digest=sha256:2fb1aac0823304bdeb57fc1211eb02dba834607af5fcd29b0c350fd0cdabb904

Observation 70f4b1c1-b9ad-4495-9d6e-c03e32f9205a · outbound

This paper cites Qwen2.5 Technical Report.

Crosslingual Reasoning through Test-Time Scaling Qwen2.5 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.429958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.429958Z digest=sha256:cd8406fab15a75b48921a66fb36059f2a7ecb1308c06d7898cec9cb92966c3f7

Observation 1e052f73-eaf9-4c51-9c9b-09a2fc4c5a9c · outbound

This paper cites Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning.

Crosslingual Reasoning through Test-Time Scaling Linguistic Generalizability of Test-Time Scaling in Mathematical Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.434873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.434873Z digest=sha256:bf6ec3c9f3248e4c66e44b4036d2941f0f475634b0ff09fb28474ad5cbe3f23c

Observation 0764cfe7-825c-41d8-84e4-986d4fff323c · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Crosslingual Reasoning through Test-Time Scaling Chain-of-thought prompting elicits reasoning in large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.440062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.440062Z digest=sha256:280345b863ac41985153ad02c718d9b55df7ad1f04c553437fe2c62aec1417f9

Observation df3335de-51ba-4e48-98a9-8904eb792d36 · outbound

This paper cites Program induction by rationale generation: Learning to solve and explain algebraic word problems.

Crosslingual Reasoning through Test-Time Scaling Program induction by rationale generation: Learning to solve and explain algebraic word problems

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.444811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.444811Z digest=sha256:d383134b7c7e2eddb11ca8c15933ef0769d9b029f88f88be1df26a4274df4f01

Observation 7f438d41-d641-41bd-ad7f-452e52249539 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Crosslingual Reasoning through Test-Time Scaling Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.455968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.455968Z digest=sha256:06b16972a6afcdd7e97047934eb19a87f80c7a4eac64ebf201889a20ac1d71e7

Observation eeb25427-bcab-4dfd-bfc8-d669d59d10af · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Crosslingual Reasoning through Test-Time Scaling Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.460959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.460959Z digest=sha256:e7f6e1eaf9ef7a9e6e021dde678c8236530d18675fb49b7b9b701f65cee357ce

Observation 0520dd9c-e99a-4ff0-93f4-fcf68b5f0946 · outbound

This paper cites Chain of thought empowers transformers to solve inherently serial problems.

Crosslingual Reasoning through Test-Time Scaling Chain of thought empowers transformers to solve inherently serial problems

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.943879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.465739Z digest=sha256:74ce2dff4fe1530ef4b1293aceb1e2847d3886dcebdbe4bd3bd14574eb4bf78d

Observation 74e7799f-e0b8-4908-8b36-980dac921ce8 · outbound

This paper cites Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs.

Crosslingual Reasoning through Test-Time Scaling Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.470291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.470291Z digest=sha256:25dd3e59605afe46749bd0f85f2222f068601b77e5f42ffe2fdd8ab0469f3a88

Observation 72fc66bc-3a79-42d0-94a3-1ef194fbea7a · outbound

This paper cites Evolving Deeper LLM Thinking.

Crosslingual Reasoning through Test-Time Scaling Evolving Deeper LLM Thinking

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.475080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.475080Z digest=sha256:c3c8b47bbcef6aeef27f9d848721831dabe1c66ab00b67b0a4a9f01d1d8f2141

Observation fa15ad93-73f0-469b-887e-3b1abbcd8c2b · outbound

This paper cites O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?.

Crosslingual Reasoning through Test-Time Scaling O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.481057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.481057Z digest=sha256:ed0912ec8764f4fab58c021d00d7386f2ea43bdaa7931dac44b02cf0f7b10554

Observation b057238c-9ad2-4fb5-b571-500180286167 · outbound

This paper cites Bespoke-stratos: The unreasonable effectiveness of reasoning distilla- tion.

Crosslingual Reasoning through Test-Time Scaling Bespoke-stratos: The unreasonable effectiveness of reasoning distilla- tion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.486032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.486032Z digest=sha256:f8b8334877405141300aaed8da6b024868e82adac4a1d2479172bbfc6d57e932

Observation 5b3a0392-1a53-405f-b3a7-d9be978e2a19 · outbound

This paper cites Millions scale dataset distilled from r1-32b.

Crosslingual Reasoning through Test-Time Scaling Millions scale dataset distilled from r1-32b

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.490717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.490717Z digest=sha256:ab9f9e4824a64a75a9bd1c56089f87638a380b91915b426651e4de7378a50558

Observation 6cadb4e2-c438-42dc-9f89-e28237972b87 · outbound

This paper cites Tina: Tiny reasoning models via lora.

Crosslingual Reasoning through Test-Time Scaling Tina: Tiny reasoning models via lora

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.905542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.495354Z digest=sha256:4b6e8218aec67a19c58130638e1713390767408a1f68c9c790a422f0182842b4

Observation b02139bd-10b4-4d54-acaa-d220ef10d13c · outbound

This paper cites The Llama 3 Herd of Models.

Crosslingual Reasoning through Test-Time Scaling The Llama 3 Herd of Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.499987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.499987Z digest=sha256:50e3a77ab92146b5e6342463715eec1fcdc3b4a304a987a62d230a0e0e783337

Observation 2f5da101-1b94-4f59-a6bd-8235f9e17842 · outbound

This paper cites Think before you speak: Training language models with pause tokens.

Crosslingual Reasoning through Test-Time Scaling Think before you speak: Training language models with pause tokens

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.504329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.504329Z digest=sha256:4b15550b650581442ce7ab2c310828c2bf87de546f0853aa6230655c5bb0d040

Observation 5eed85ec-2ebc-486e-944e-dcdd8120cc31 · outbound

This paper cites Language models are multilingual chain-of-thought reasoners.

Crosslingual Reasoning through Test-Time Scaling Language models are multilingual chain-of-thought reasoners

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.878920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.509040Z digest=sha256:5e3bdec1a4d89f6267b7be0965815b9a4fdb7948b6f5257cbd961e7d2fee5516

Observation ecc9e832-c3e0-40be-a3de-3fbc3672e6d4 · outbound

This paper cites Cross-lingual prompting: Improving zero-shot chain-of-thought reasoning across languages.

Crosslingual Reasoning through Test-Time Scaling Cross-lingual prompting: Improving zero-shot chain-of-thought reasoning across languages

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.862182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.513637Z digest=sha256:2d939612dfa3f75248e6c19e857b3ae1626d2378e3b8058f1b970d7374a0ecf3

Observation 3dd01407-6afd-4be5-8c2f-e989d9376cb4 · outbound

This paper cites Question translation training for better multilingual reasoning.

Crosslingual Reasoning through Test-Time Scaling Question translation training for better multilingual reasoning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.845065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.518466Z digest=sha256:59316fdd119e43df5275224ac0b812e40275a407f506d1abc827c6d93ac63768

Observation 2acf8cb6-1bdd-46de-a354-730ed2075c48 · outbound

This paper cites Understand, Solve and Translate: Bridging the Multilingual Mathematical Reasoning Gap.

Crosslingual Reasoning through Test-Time Scaling Understand, Solve and Translate: Bridging the Multilingual Mathematical Reasoning Gap

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.523233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.523233Z digest=sha256:f7d67fad633e5c9897d02444cbf3f9bb48f698d37093288bec6a9de490d7f7f3

Observation 3b6afc4f-3825-4aa3-8c26-752fcd3cae8f · outbound

This paper cites LangBridge: Multilingual reasoning without multilingual supervision.

Crosslingual Reasoning through Test-Time Scaling LangBridge: Multilingual reasoning without multilingual supervision

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.827441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.528607Z digest=sha256:ac398fb01dc81a25c242fe48e44f4aeafbf72a1c2bde97e627152d1194b6edcc

Observation fece9867-8a65-4a66-8fef-7317f7e906ff · outbound

This paper cites Mindmerger: Efficiently boosting LLM reasoning in non-english languages.

Crosslingual Reasoning through Test-Time Scaling Mindmerger: Efficiently boosting LLM reasoning in non-english languages

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.810692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.533615Z digest=sha256:6c2fdf8c0880db377e2b11589610d74633c6235899c2f0b66b0e704ca67074ba

Observation f9a32539-2947-446b-b6bb-74ab7c389de4 · outbound

This paper cites MAPO: Advancing multilingual reasoning through multilingual-alignment-as-preference optimization.

Crosslingual Reasoning through Test-Time Scaling MAPO: Advancing multilingual reasoning through multilingual-alignment-as-preference optimization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.794602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.538643Z digest=sha256:ba81858c4739a8d0b7f0cfa17dd84faf721200bce785cd14c99e82a4173ea220

Observation 81707aae-6fd5-4846-8823-a226c1df4b1a · outbound

This paper cites Language imbalance driven rewarding for multilingual self-improving.

Crosslingual Reasoning through Test-Time Scaling Language imbalance driven rewarding for multilingual self-improving

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.778277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.543521Z digest=sha256:6ac004a6a2d6babf963be11d0f8c0e4a1e39ebb6b891d86da8d3771b2f5501d6

Observation e0b07b3b-4ee4-4fc9-82cf-da826fdda5f2 · outbound

This paper cites Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Observations.

Crosslingual Reasoning through Test-Time Scaling Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Observations

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.548452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.548452Z digest=sha256:895fac1035b41ba9d2cd0e7bedc94c19e51cda186627cfba8305656c70f047a4

Observation fe1787ce-433b-43eb-b9b2-f4fe6da952de · outbound

This paper cites Language model developers should report train-test overlap.

Crosslingual Reasoning through Test-Time Scaling Language model developers should report train-test overlap

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.553488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.553488Z digest=sha256:b0d5148bf95ac09dbe7cbc0c2e27d92ebccb5e39b06cff2ebf1a55ed01b66d5f

Observation 72db4342-d5fb-488e-a255-0b6f62d2810d · outbound

This paper cites an unresolved cited work.

Crosslingual Reasoning through Test-Time Scaling Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T23:08:01.762396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.558933Z digest=sha256:7119df09caf92ba306c0a3d80bf5aed55e95be6e48f315b7cef02b1f2e982df5

Observation 185b32b3-61be-4cff-9e80-05ce8c03b454 · outbound

This paper cites Measuring massive multitask language understanding.

Crosslingual Reasoning through Test-Time Scaling Measuring massive multitask language understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.563492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.563492Z digest=sha256:cc86236373e7d377b635a788158f0ae186456a0e2ac64a03af79051a3c87aec2

Observation a8be1b13-1a4e-487d-9564-4d1864df9684 · outbound

This paper cites FORK: A bite-sized test set for probing culinary cultural biases in commonsense reasoning models.

Crosslingual Reasoning through Test-Time Scaling FORK: A bite-sized test set for probing culinary cultural biases in commonsense reasoning models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.735254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.568207Z digest=sha256:0d546c2427afc802f36e86df8a15705e763ae1469680423555cfb2c026489610

Observation ab505657-b99c-4ee2-a989-9bcae0f1dfbb · outbound

This paper cites CommonsenseQA: A question answering challenge targeting commonsense knowledge.

Crosslingual Reasoning through Test-Time Scaling CommonsenseQA: A question answering challenge targeting commonsense knowledge

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.572782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.572782Z digest=sha256:f9d91b7fbb022c5173a694934b2b3e6833c466af11885fe94a8e819991d3f515

Observation 0d39e314-adbe-498f-9b61-d87cb78545d8 · outbound

This paper cites COPAL- ID: Indonesian language reasoning with local culture and nuances.

Crosslingual Reasoning through Test-Time Scaling COPAL- ID: Indonesian language reasoning with local culture and nuances

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.708348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.577263Z digest=sha256:0481ca42a922d654c45d0a31bdbebf1606d554c1a57a8031c0c963bea0d6ff9d

Observation 616e5740-c45a-46cd-9475-9695bff95a60 · outbound

This paper cites Choice of plausible alterna- tives: An evaluation of commonsense causal reasoning.

Crosslingual Reasoning through Test-Time Scaling Choice of plausible alterna- tives: An evaluation of commonsense causal reasoning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.691708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.581780Z digest=sha256:e1ac43829f3411de1e264152870a60cd38d70b7a9d5b890bec6b4330eb205319

Observation 0f260dfe-0f1f-4c6e-b68a-e871d71dd550 · outbound

This paper cites A framework for few-shot language model evaluation, 07 2024.

Crosslingual Reasoning through Test-Time Scaling A framework for few-shot language model evaluation, 07 2024

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.586668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.586668Z digest=sha256:9f64d577193842f3a001a8b9a5ae68a9ba788c5157e500abef59858ff5b9aaaf

Observation 769e99e8-af4f-44d6-bd5a-d01074ffc475 · outbound

This paper cites MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning.

Crosslingual Reasoning through Test-Time Scaling MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.591261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.591261Z digest=sha256:ea6c0cb196f04fd5a048d1b9a5cb0e16ecfc337f928da5c593e55c7fe7815c29

Observation 949d3b87-c155-4382-a776-7d7dede4d523 · outbound

This paper cites SLAM: Towards efficient multilingual reasoning via selective language alignment.

Crosslingual Reasoning through Test-Time Scaling SLAM: Towards efficient multilingual reasoning via selective language alignment

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.664871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.596599Z digest=sha256:337ed53487d21605ff4297b53fc0f3848cd0a29e30d714b1a0a305fbd3e5fff7

Observation 1cc88bbf-b27c-41fe-872e-73767d91d967 · outbound

This paper cites Gemma 3 Technical Report.

Crosslingual Reasoning through Test-Time Scaling Gemma 3 Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.601436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.601436Z digest=sha256:5b8fd839a21bb9b961a91f78aaab7b3d873fb38dccf4e450a64bd59e275a935d

Observation ce9e398f-8114-404d-88c2-6813e20b1fa7 · outbound

This paper cites Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws.

Crosslingual Reasoning through Test-Time Scaling Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.606321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.606321Z digest=sha256:27ebc90ef3b8e5b80548a71ed7acd71505f790dfab81bcbe185fd88f98ca9869

Observation cb08a361-f53c-420d-8642-b9f9871e77c0 · outbound

This paper cites Metamath: Bootstrap your own mathematical questions for large language models.

Crosslingual Reasoning through Test-Time Scaling Metamath: Bootstrap your own mathematical questions for large language models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.647036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.611284Z digest=sha256:355c83a18a2c0bafdc735741b7db36229a1c0493162426268ccfd7f83d53bf6a

Observation 64083c27-407d-4779-bfc2-bb9c0086eb43 · outbound

This paper cites Qwen3: Think deeper, act faster, 4 2025.

Crosslingual Reasoning through Test-Time Scaling Qwen3: Think deeper, act faster, 4 2025

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.631083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.616203Z digest=sha256:d8218fbf96861b063122e3577bffb20ccb763708b3a3c377de34652bc4fe8356

Observation 613a44bb-b0e7-4883-a6c9-2d3d25919089 · outbound

This paper cites Foreign-language quotations and code-switching: The grammar behind.

Crosslingual Reasoning through Test-Time Scaling Foreign-language quotations and code-switching: The grammar behind

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.614988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.621320Z digest=sha256:762210bc9e45d3d4fb418dfd6128b8d902e488aa54002c67f2f44031e2c2dd7f

Observation 1518f66e-d0a6-4e5e-bb31-ffe224b518da · outbound

This paper cites The state and fate of linguistic diversity and inclusion in the NLP world.

Crosslingual Reasoning through Test-Time Scaling The state and fate of linguistic diversity and inclusion in the NLP world

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.599461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.626236Z digest=sha256:a9b8a702203f30040a38102a4ed16dd1c4e95519f4755b28c8ad4ae113ea4e74

Observation 336c80b5-b12d-4e72-a5e9-3ad088fc9918 · outbound

This paper cites Language model tokenizers introduce unfairness between languages.

Crosslingual Reasoning through Test-Time Scaling Language model tokenizers introduce unfairness between languages

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.584409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.630751Z digest=sha256:9c1321cd74424dfed576bdd64edab5cb7b6e65476b4d32a42e25a72cfa0d2f48

Observation 143eff3b-f4e6-4db1-85ce-e629ab29ad04 · outbound

This paper cites Smith, and Yulia Tsvetkov.

Crosslingual Reasoning through Test-Time Scaling Smith, and Yulia Tsvetkov

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.568043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.635351Z digest=sha256:6c08e8953f82dd08d034f2bfed34a922cf60835e8bce6c8c7c52807a80f3bd6c

Observation b1413cf4-9b8f-42f0-a60d-3fe98aa68ba8 · outbound

This paper cites Olympicarena: Benchmarking multi-discipline cognitive reasoning for superintelligent ai.

Crosslingual Reasoning through Test-Time Scaling Olympicarena: Benchmarking multi-discipline cognitive reasoning for superintelligent ai

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.552753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.639979Z digest=sha256:e881fdae33b02c7c88208eeec63fc37fa7581bf0155171e7951c0cfbe016c062

Observation 5b530b99-ed81-41d9-b7d7-a6ec45dcb8d8 · outbound

This paper cites Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse.

Crosslingual Reasoning through Test-Time Scaling Mind Your Step (by Step): Chain-of-Thought can Reduce Performance on Tasks where Thinking Makes Humans Worse

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.644773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.644773Z digest=sha256:450e7e36407efc72dfed56ef9f522d59c3c2ac268a42dd9c7920f7d526aab7ff

Observation ba977ab3-9ac9-4241-8d62-e1baf8f0fd33 · outbound

This paper cites The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks.

Crosslingual Reasoning through Test-Time Scaling The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.649709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.649709Z digest=sha256:4a783d80adc17a242c20b0f75c4ad60ca7d067752ed6d10b8ba5303cd16cb9c7

Observation 831a1e3b-f8df-4b89-bbd6-b030c4606009 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Crosslingual Reasoning through Test-Time Scaling Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.654610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.654610Z digest=sha256:1708c5633ed0c0379c592bdbf2efc9c4a72458c2504ab9ec672c875cde4690f6

Observation 36d42f51-9370-4da3-bebe-97026403fde7 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

Crosslingual Reasoning through Test-Time Scaling Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.659695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.659695Z digest=sha256:f25d3f7e48739ea3d82af1bf870321a8e95742c69419197b6f83cafe04f6ae59

Observation c3906f4b-4903-487c-9188-902919c3c947 · outbound

This paper cites BLOOM+1: Adding language support to BLOOM for zero-shot prompting.

Crosslingual Reasoning through Test-Time Scaling BLOOM+1: Adding language support to BLOOM for zero-shot prompting

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.537371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.664573Z digest=sha256:684eb941530b9abf6519fc6a4ad7e882ec4d8e47a49441038f58606f30eabed1

Observation c868e812-d26b-41a8-bf33-267c891858ff · outbound

This paper cites Understanding catastrophic forgetting in language models via implicit inference.

Crosslingual Reasoning through Test-Time Scaling Understanding catastrophic forgetting in language models via implicit inference

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.521886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.669314Z digest=sha256:c3f724ae845b4b5beb598e7781ff25949bae87b0fb68409ed46479e39bb549d2

Observation d95a6dff-0406-43a5-8e27-8b24c5d4c9cd · outbound

This paper cites Translating across cultures: LLMs for intralingual cultural adaptation.

Crosslingual Reasoning through Test-Time Scaling Translating across cultures: LLMs for intralingual cultural adaptation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.506553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.674024Z digest=sha256:66c96656167cca446b81a8374a88b7e19aaf713fe8927750d8422f5657ded147

Observation c65ba839-b9c2-4c1b-b2ba-82e863274481 · outbound

This paper cites Multilingual != Multicultural: Evaluating Gaps Between Multilingual Capabilities and Cultural Alignment in LLMs.

Crosslingual Reasoning through Test-Time Scaling Multilingual != Multicultural: Evaluating Gaps Between Multilingual Capabilities and Cultural Alignment in LLMs

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.678858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.678858Z digest=sha256:14056f737f1f1b9f46be27af4f08afeef936030a05851565a39f9b05578f5f55

Observation 863c91d6-4ca6-4aab-8506-6e7790aec02d · outbound

This paper cites Mortensen, and Graham Neubig.

Crosslingual Reasoning through Test-Time Scaling Mortensen, and Graham Neubig

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.490695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.683823Z digest=sha256:ed37f4fb869002d08990cc9a97b8c89ca620d89ac74117b58c55aadf6a8d33ac

Observation b0524281-23f6-4547-b6a1-b75f5179fb92 · outbound

This paper cites Shortcomings of LLMs for low-resource translation: Retrieval and understanding are both the problem.

Crosslingual Reasoning through Test-Time Scaling Shortcomings of LLMs for low-resource translation: Retrieval and understanding are both the problem

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.474661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.688540Z digest=sha256:cf9396cb4bbd0ab96216d636695b0d49af455740a3077d93a19093c458f11ee6

Observation bddd0366-8877-452f-bad0-664f06acac78 · outbound

This paper cites Is Small Language Model the Silver Bullet to Low-Resource Languages Machine Translation?.

Crosslingual Reasoning through Test-Time Scaling Is Small Language Model the Silver Bullet to Low-Resource Languages Machine Translation?

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.693181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.693181Z digest=sha256:635afb12b9fc80839a375e175deab5d599e267cb88c5f8b1108f86492d6e2503

Observation af0e3dd4-c2b3-48ec-bc9c-4571819a52ea · outbound

This paper cites LLM-powered data augmen- tation for enhanced cross-lingual performance.

Crosslingual Reasoning through Test-Time Scaling LLM-powered data augmen- tation for enhanced cross-lingual performance

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.458061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.697986Z digest=sha256:b0dc7a52a9f048b31abfd04840e701a02df08af40180e99141fe7be724430fda

Observation 56ef7306-cc47-42f5-a8ab-083b159b8b0a · outbound

This paper cites LexC-gen: Generating data for extremely low-resource languages with large language models and bilingual lexicons.

Crosslingual Reasoning through Test-Time Scaling LexC-gen: Generating data for extremely low-resource languages with large language models and bilingual lexicons

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.440991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.702497Z digest=sha256:81ce2c6b456e9c44dc8e14671b321bc7a64fc2d30a7016ac51f86ac66a10b23e

Observation 867c3e83-1c35-4898-afd1-7d378d25396e · outbound

This paper cites XLM-V: Overcoming the vocabulary bottleneck in multilingual masked language models.

Crosslingual Reasoning through Test-Time Scaling XLM-V: Overcoming the vocabulary bottleneck in multilingual masked language models

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.425549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.707053Z digest=sha256:af6003bc5e77cb888ab5fe48b1a937f756c31c59a27daf36064f6b884d05f7d2

Observation 2cd128f8-a3ca-4668-a5e5-1c9550de1247 · outbound

This paper cites Adapters for altering LLM vocabularies: What languages benefit the most? In The Thirteenth International Conference on Learning Representations, 2025.

Crosslingual Reasoning through Test-Time Scaling Adapters for altering LLM vocabularies: What languages benefit the most? In The Thirteenth International Conference on Learning Representations, 2025

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.408281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.712021Z digest=sha256:6f1b2c9dd064967f6fb046846f568c32e456bf59cc4a08fb5a72d45bf8cdb775

Observation 7adc6218-2a17-48dc-8834-b528ca8de777 · outbound

This paper cites ByT5: Towards a token-free future with pre-trained byte-to-byte models.

Crosslingual Reasoning through Test-Time Scaling ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T23:08:01.391304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.716567Z digest=sha256:5f674495281a1ee59f7ade6f430a9d59e66aff0b425d4fb9d1f8cd719e189022

Observation 9c0e533e-3d22-4e0e-b018-c527f898084a · outbound

This paper cites Small models struggle to learn from strong reasoners.

Crosslingual Reasoning through Test-Time Scaling Small models struggle to learn from strong reasoners

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.721221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.721221Z digest=sha256:0296d21a98408b437d1579390b501483fa7b028da41ae329157fa3198d9524bd

Observation 6d5f2e68-e5cb-417b-8c00-608a8fa2ccea · outbound

This paper cites How Do Multilingual Language Models Remember Facts?.

Crosslingual Reasoning through Test-Time Scaling How Do Multilingual Language Models Remember Facts?

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.725796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.725796Z digest=sha256:a47e3be72c151af71859ed3daa7deae96c5e5536328abdb6c0a634aa30f673db

Observation 10ada909-1c14-4ea0-a304-1acd3200b3a5 · outbound

This paper cites Sie isst 3 Eier zum Frühstück und verwendet 4 Eier für Muffins, also verwendet sie insgesamt 3 + 4 = 7 Eier pro Tag.

Crosslingual Reasoning through Test-Time Scaling Sie isst 3 Eier zum Frühstück und verwendet 4 Eier für Muffins, also verwendet sie insgesamt 3 + 4 = 7 Eier pro Tag

Reference 77

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T23:08:01.375290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T23:08:00.731058Z digest=sha256:910e55a999882371b3c2ff00746242cb03cecc83a1ed1c0fcb0dc74039194cbc

Observation d6de4d60-f612-40c2-b0cc-c3f8b41c2938 · outbound

This paper cites an unresolved cited work.

Crosslingual Reasoning through Test-Time Scaling Unresolved cited work

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.450740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.450740Z digest=sha256:083e01d64d7b665b5ee533e1335431264ff52b4452b2b69de240b67327e89651

Pith citing papers

Observation bbb98af1-5c7c-4f69-a858-41de547fd90f · inbound

Language Specific Knowledge: Do Models Know Better in X than in English? cites this paper.

Language Specific Knowledge: Do Models Know Better in X than in English? Crosslingual Reasoning through Test-Time Scaling

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:51:42.702671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T14:48:29.590965Z digest=sha256:75c273eda4e23492aee723b8fdfd6e578be2aecf8b21c0cf38c07c6f04d05768

Observation c6a56755-adc8-461d-b385-43969dc9683d · inbound

ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark cites this paper.

ARB: A Comprehensive Arabic Multimodal Reasoning Benchmark Crosslingual Reasoning through Test-Time Scaling

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:27.736254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:27.736254Z digest=sha256:c5f77183a35105aef6ac5cf9519b197313a6685b84874d26275919e7258f3bd4

Observation f929a16c-a01e-47f1-8449-ec7209d572a0 · inbound

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It cites this paper.

The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It Crosslingual Reasoning through Test-Time Scaling

Reference 140

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:28.238008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:40:28.238008Z digest=sha256:1f8d1e022b35b42f55d3a7dcf0b7024e4ac2212222fbfcf548625c2556133518

Observation dd5f3866-04b2-49e0-a982-4031a41ee8ad · inbound

Language as a Latent Variable for Reasoning Optimization cites this paper.

Language as a Latent Variable for Reasoning Optimization Crosslingual Reasoning through Test-Time Scaling

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:31:06.570780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-09T21:40:37.499246Z digest=sha256:63771c6fbd52528643d01689d03f9bb7c12b9c703ec4852a33f4180d8558b609

Observation baab054c-e75c-4ba4-b9ad-c4b84b9c1ac5 · inbound

Prompting language influences diagnostic reasoning and accuracy of large language models cites this paper.

Prompting language influences diagnostic reasoning and accuracy of large language models Crosslingual Reasoning through Test-Time Scaling

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:28:12.479499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-20T10:24:30.749958Z digest=sha256:48ceef078af5f050b6f8b25244739c2cc42f5f5ed356a908afd53d479c45d995

Observation 30689f68-2d9f-4fe7-981c-5de38d68dd74 · inbound

LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance cites this paper.

LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance Crosslingual Reasoning through Test-Time Scaling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:21:09.735534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-22T06:19:44.377733Z digest=sha256:0fb6edafb1629a30fa8197c1f768ffc2aa668d0d2dceba3a2016c824ddd2c1f2

Observation aae947fd-b0ca-46a6-9cc6-a545f606909d · inbound

Soft Token Alignment for Cross-Lingual Reasoning cites this paper.

Soft Token Alignment for Cross-Lingual Reasoning Crosslingual Reasoning through Test-Time Scaling

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:52.109759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-26T05:43:20.437278Z digest=sha256:2fc8109c98d4ebdcf0e1bb09f8a2e566eb52826ccbf8cd9d6eea416b2a7e9ffd

Observation 7a8bee9a-4abb-4eb7-ba52-9cd2ea191ff0 · inbound

Cost of Reasoning in non-English Languages: A Case Study on Japanese cites this paper.

Cost of Reasoning in non-English Languages: A Case Study on Japanese Crosslingual Reasoning through Test-Time Scaling

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T14:12:44.286699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T14:12:44.286699Z digest=sha256:b8a1ea9681cbaed9c68f919a05a1486a6d7c1105a7d51d1a9f08d41f171f3418