Pith. sign in

Paper Citation Record · LEDGER

Magistral

As of 9 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 40 inbound Pith citation observations for arXiv:2506.10910.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.10910 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:19:26.848892Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:11:01.995426Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 58df91f7-f6d0-43f2-914c-f803515f43e7 · outbound

This paper cites What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study.

Magistral What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:22.151310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:22.151310Z digest=sha256:df5a0ded52aac656333ab6a7ebe0eeff3be5471d5640554e5ccffa697bed21fc

Observation 5c547a00-d3e3-4bef-9df6-0750e017b308 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Magistral DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:22.289997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:22.289997Z digest=sha256:c3df4ae8f1697616a3bc475563fb94356aae5b3d124bc565c5f54545e0fbd5ac

Observation 5704b324-819e-4fe2-bbb4-d084dc1041c6 · outbound

This paper cites Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures.

Magistral Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:29.137514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:22.472941Z digest=sha256:3c42fb2f552864e808171d693721ea11ab6967d60f46abd7fadf66d35a86c09f

Observation 67e25e8a-2b82-49b5-921b-5f515a8f23fb · outbound

This paper cites Polyglot Benchmark.

Magistral Polyglot Benchmark

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:28.977000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:22.624332Z digest=sha256:cf6d7641b023220bcbaf37a6e8bc88cb563f28e48878a385551d5bd7353e405e

Observation 5003029f-9c06-4414-9506-4461b56d06f7 · outbound

This paper cites OpenThoughts: Data Recipes for Reasoning Models.

Magistral OpenThoughts: Data Recipes for Reasoning Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:22.778940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:22.778940Z digest=sha256:b98d2e8cc1ca199893a08c9ec0f4ee6350236324075b7e8f5e2dc9f89a0a5bf0

Observation 68a4518f-83be-443b-9663-c4830f70049d · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Magistral Measuring Mathematical Problem Solving With the MATH Dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:22.981703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:22.981703Z digest=sha256:9e1800cefed97a1c577a484cb9ab2430e1845255a5e8d54c38c391557c34b2d2

Observation 8d724e63-8a2b-4c53-9039-7a7c68ed16e7 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

Magistral OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.130127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.130127Z digest=sha256:0587d7b7058b2d6a2737f920964ea59ca3ce737c420674924f36276cc48aa342

Observation 3aaefb17-6335-4a84-9917-cbe598a1b973 · outbound

This paper cites Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model.

Magistral Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.278363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.278363Z digest=sha256:b6f694655414fcf238e5df43e851d75aac1d1d8111915f8df0cf918da2c5bf69

Observation cbf1b255-9e5c-4119-80eb-1f58bd63ee53 · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

Magistral Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.479130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.479130Z digest=sha256:36ac24ba2b98892cd664bb53b3ad4f33030a0106ab5b5c3774e236b42befa0a2

Observation 48ba2a9f-8a07-4070-8b93-7762621a965a · outbound

This paper cites OpenAI o1 System Card.

Magistral OpenAI o1 System Card

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.630431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.630431Z digest=sha256:d7fb702a22487518100e98b94e52c329a6513a5616b380492d667322454bf8d7

Observation 0b17f5c1-93e3-476c-9a23-20a2d86b9f3f · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Magistral LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.789584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.789584Z digest=sha256:c468db904aabad4d3555dd7d6e8e62047ec745c4852a5f6d53de38d1c97ccac8

Observation f778baa5-f855-40d6-b5b6-d2f99d9e1d71 · outbound

This paper cites FastText.zip: Compressing text classification models.

Magistral FastText.zip: Compressing text classification models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:23.974974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:23.974974Z digest=sha256:644fbf569ef6de9f918303d36228b92d7617335e6a20571f3ce3c0275b99ec62

Observation fec18fcf-db8d-4007-8074-085886f07043 · outbound

This paper cites Visualizing the Loss Landscape of Neural Nets.

Magistral Visualizing the Loss Landscape of Neural Nets

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:24.080434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:24.080434Z digest=sha256:a06a5c07e4f397d5ee43747c0467b4b111518784a3d6e0ef5844bc0e53488880

Observation b9445add-fb95-4f97-8b48-afd2089ea90f · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Magistral Understanding R1-Zero-Like Training: A Critical Perspective

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:24.249365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:24.249365Z digest=sha256:2170a987d388cc63b8d494df37cc1b99b93cd8ce296ab791d3d586f218c4af57

Observation 628e55f3-111d-414e-b23d-9926e437097c · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts.

Magistral Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:24.433838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:24.433838Z digest=sha256:268f7ceac10c4a5d5eaf8faa45bbde030e683b561f50fdfa0ff635794fa6720d

Observation b9cbcabf-32ce-4767-bc51-ebc7e5781d15 · outbound

This paper cites Mistral large 2.

Magistral Mistral large 2

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:28.649078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:24.559640Z digest=sha256:22d153655797c2fca04354c775324c86ae287fc8faa323099ead8506a039d5c0

Observation ba51863e-03b3-4038-9708-26edd0658d47 · outbound

This paper cites Mistral medium 3.

Magistral Mistral medium 3

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:28.343947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:24.717927Z digest=sha256:65a7a7e5133720eb3ce6f85d2de94230b56075985297cf4bfad46f1f70ef8c9f

Observation 80499156-04c0-4aec-b60a-9c656f3b4258 · outbound

This paper cites Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models.

Magistral Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:24.817910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:24.817910Z digest=sha256:43ee0c99ffe7b3d364ec3ca14857509d05ce39b39a2821434602c4dbbc24008f

Observation 24591ab7-adab-4bca-adea-f95306197332 · outbound

This paper cites Codeforces cots.

Magistral Codeforces cots

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:28.009074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:24.973926Z digest=sha256:035c521394ac7ddfedcb8a126daab294f6ea7d35553a6c723faccf17dc8ae1a2

Observation 4c805a10-7964-4e0b-9367-b218e3a0d53c · outbound

This paper cites Humanity's Last Exam.

Magistral Humanity's Last Exam

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:25.101520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:25.101520Z digest=sha256:60622b62ffb2db754ff44acb7318699fc2eee431e9e2699f15b6213ea1d53b0a

Observation 21d0f88c-bac0-4a20-ba78-ad10edf48079 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

Magistral Gpqa: A graduate-level google-proof q&a benchmark

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:25.209936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:25.209936Z digest=sha256:117a9ffbda3866621c508f37752bef97e5ab9dbe154cc79ee28d8dc006d88110

Observation 948f4ce8-b893-48e3-a349-082b1c6adf57 · outbound

This paper cites Iterative methods for sparse linear systems.

Magistral Iterative methods for sparse linear systems

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:27.763019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:25.393789Z digest=sha256:6d61583fb8829868d94f07a26c6f36cdca407ef09427ae537818732c7235a386

Observation d77bcf82-acb8-45a3-8aff-7a87174bf9c7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Magistral Proximal Policy Optimization Algorithms

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:25.552219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:25.552219Z digest=sha256:d561a81cdd02b4dd6063d083eb89c5494a83d6d0cf0068ce30bd0c1c655f7b9f

Observation 6a9d1a23-b828-4223-af6c-82725fa9ac0b · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Magistral DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:25.739088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:25.739088Z digest=sha256:b56136519f4e079de760a4aaab200350471bdbaa014b27954848bb879cdf09f1

Observation 7f4314a9-ef00-41f8-afec-0260a50552f4 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Magistral HybridFlow: A Flexible and Efficient RLHF Framework

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:25.902064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:25.902064Z digest=sha256:a28f03fc7920112471eaa576804c253c95b4d46d5870942cd7092767fda3f394

Observation b0e1489a-96a4-46bd-90f8-18784a3d8458 · outbound

This paper cites Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning.

Magistral Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:26.065118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:26.065118Z digest=sha256:3ca43f630d57879abe6be09060ad1fd32fffe07d7b9ba1fd8f265aeec440ce8e

Observation 62aa2849-4bc7-4bce-953f-b36661f1c2ce · outbound

This paper cites LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training.

Magistral LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:26.241394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:26.241394Z digest=sha256:d323bdd0d3513e7fa98a8d1633a33266d4caf6bcf83a1018cd5b590383f61f16

Observation df998082-8d39-4586-bd26-e31f6501b6fd · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Magistral DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:26.415218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:26.415218Z digest=sha256:88e5560a077c725233c2c231430c9e3dc4ab705e91bbcbd3df3929db8ff778f7

Observation afb421ca-e8f2-4607-8461-078a57556093 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

Magistral Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:26.563646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:26.563646Z digest=sha256:49364f104e3741c469afcfd2186e5956520deff2c4cd243417afe973c393b506

Observation 7f2c72e9-0ff2-4342-a537-1c5365d8db3d · outbound

This paper cites Mmmu-pro: A more robust multi-discipline multimodal understanding benchmark.

Magistral Mmmu-pro: A more robust multi-discipline multimodal understanding benchmark

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:19:27.582378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:19:26.723792Z digest=sha256:875c03e1b859cccdd46454c109ccb37526520f02768140d7eff0003f0505631d

Observation 52cca6b6-2d53-4db9-9b01-09a147b99b98 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Magistral Instruction-Following Evaluation for Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:19:26.848892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:19:26.848892Z digest=sha256:c80c802cc6d76ee6bd383b9ae2fe43c4119aa161efa92b0fb6bb0b54f22181f2

Pith citing papers

Observation 55adf272-4e07-4ef0-975f-904dc0245558 · inbound

MiroMind-M1: An Open-Source Advancement in Mathematical Reasoning via Context-Aware Multi-Stage Policy Optimization cites this paper.

MiroMind-M1: An Open-Source Advancement in Mathematical Reasoning via Context-Aware Multi-Stage Policy Optimization Magistral

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:55:42.560228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:55:42.560228Z digest=sha256:01a1b6002b179e73e5921e7ade9a27fae1bc58d784687d53db53660c890d331d

Observation 47fa10e6-114b-4b08-b746-7b13c998cd2c · inbound

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning cites this paper.

Can One Domain Help Others? A Data-Centric Study on Multi-Domain Reasoning via Reinforcement Learning Magistral

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T14:53:04.564942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:53:04.564942Z digest=sha256:8e5dea12cd70f3bde7e5a93b4d6988b12e9dda0ca6b4f055558174be7eff887b

Observation b779c2bb-1bff-4a9e-b9c9-72297ed03027 · inbound

Flow Matching Policy Gradients cites this paper.

Flow Matching Policy Gradients Magistral

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:09.588485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:09.588485Z digest=sha256:8b70a146a3f68cb454b72b322a05a0f38b954506287e1348edc3ec22d5cfa801

Observation 7edf36cd-5d18-4ee0-97ca-7ebc8605bc18 · inbound

Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency cites this paper.

Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency Magistral

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T01:02:46.136737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:02:46.136737Z digest=sha256:166b3361dcf5013631950fb1c811a1e37f1a6c5cb398da1a40f030c161bd91a2

Observation 45628e2c-bc46-4f8f-9750-044521932656 · inbound

Hermes 4 Technical Report cites this paper.

Hermes 4 Technical Report Magistral

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T16:32:54.579019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:32:54.579019Z digest=sha256:d651736fbe718400dc635d6bcf769301ff809aba361d640b26e2ecb7afbaf81c

Observation a6555b84-43a2-475f-b906-cd989fed6db3 · inbound

rStar2-Agent: Agentic Reasoning Technical Report cites this paper.

rStar2-Agent: Agentic Reasoning Technical Report Magistral

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T14:57:39.459947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:57:39.459947Z digest=sha256:ff06ff48c5e750fa39a97859ae90940124d121ce946e123d4020b86d71796c64

Observation cd458219-c9f1-439e-ba58-d2f6ef8fa086 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle Magistral

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:38.099803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:38.099803Z digest=sha256:87319170e67b1314f3e825ad2af1cf9aed32e2ce21212988f5abe28906cd3e5c

Observation 6bfc45b6-5981-4322-8ba9-320755b5e74a · inbound

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts cites this paper.

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts Magistral

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T20:54:21.599958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T20:53:58.198974Z digest=sha256:e59306dffdc82bf0fda2fa8af892fea6bb258e42b3d2241f00c8b37ec18382d1

Observation f6298168-ee72-4d00-936c-212358e8e52e · inbound

Are Large Reasoning Models Interruptible? cites this paper.

Are Large Reasoning Models Interruptible? Magistral

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T10:07:52.637167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:07:52.637167Z digest=sha256:7b5135ed3cdec8f88a498c6a929227311782a988f55ea3580a522e4dedfebf60

Observation 8f14ebbe-2e79-4187-a375-14693ba94948 · inbound

Ministral 3 cites this paper.

Ministral 3 Magistral

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:12:24.787396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:12:24.627033Z digest=sha256:c58295e5fde5a222322ea7888ee0adb4772a1c12fefcface578460c9e355d048

Observation 44d95e98-01b7-446c-8e6f-a5982282210a · inbound

Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models cites this paper.

Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models Magistral

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:54:31.109524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:54:30.992080Z digest=sha256:6aad80a7596da69f2c055a353eca45277f193a2f0b6ab6e9bef6d15ab024fe24

Observation 66e62282-7e92-4464-b7fd-5efd6a4a2270 · inbound

Simultaneous Speech-to-Speech Translation Without Aligned Data cites this paper.

Simultaneous Speech-to-Speech Translation Without Aligned Data Magistral

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T00:57:01.017891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:57:01.017891Z digest=sha256:05b8bde597a4d35132eabe8ae66438de6f4ab3c5714851c53d72661cf5a09d05

Observation 7f4e80d6-e4f4-42ed-ba6b-3d5c09ab1a63 · inbound

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling cites this paper.

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling Magistral

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T23:14:39.407159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:14:39.407159Z digest=sha256:497a3a49883810fbca45566820e444ebd705ff2df7c5e9dbd6bb2ba5eb0bf9cd

Observation 2f7f2d9e-b2fd-4310-9f85-63e6be00e0ed · inbound

Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis cites this paper.

Task Complexity Matters: An Empirical Study of Reasoning in LLMs for Sentiment Analysis Magistral

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T20:07:28.015727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:07:28.015727Z digest=sha256:05705ca4c66cd33757ec806f995f76dc64c204d0cdda0ba2321339fc77164434

Observation 4c23408e-dc3c-4095-a984-bb2e6ebd51b7 · inbound

A Novel Hierarchical Multi-Agent System for Payments Using LLMs cites this paper.

A Novel Hierarchical Multi-Agent System for Payments Using LLMs Magistral

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T20:05:40.456510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:05:40.456510Z digest=sha256:476e4bc21eff5048e235f6cacfd16633437304651a1d1b7b2a9de7120d71b7d0

Observation 44806c9e-e9ac-4a7b-916d-e5f39cb499c4 · inbound

LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics cites this paper.

LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics Magistral

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T20:05:56.999962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:05:56.999962Z digest=sha256:b47fbc5815087dc740cfdbcd53492cc4e00073e3f08ce48a04bf732b636be535

Observation fee22f91-491a-47a8-8cf6-b9927dcdba70 · inbound

TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios cites this paper.

TSHA: A Benchmark for Visual Language Models in Trustworthy Safety Hazard Assessment Scenarios Magistral

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T17:05:14.494409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:05:14.494409Z digest=sha256:d2b219a295f92513840482126a12d95e316848b9d1f40af98f8baa5d82060088

Observation 8a26308b-c52a-449b-b4aa-de02d9f608a0 · inbound

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints cites this paper.

Empirical Evidence of Complexity-Induced Limits in Large Language Models on Finite Discrete State-Space Problems with Explicit Validity Constraints Magistral

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T14:15:29.394776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T14:12:45.438246Z digest=sha256:3c4744dfb85eb4d4a7bb15ad78d83b5d343957067c2b726e80cd2f326a8c8c97

Observation 54215b74-247f-4322-b231-219345a6c660 · inbound

Beyond Distribution Sharpening: The Importance of Task Rewards cites this paper.

Beyond Distribution Sharpening: The Importance of Task Rewards Magistral

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:08:27.020410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T08:07:14.691463Z digest=sha256:58da249132858f9bf440d6cdb437f2e1d6105e3968c8d8ce44b3e14ef7c120b9

Observation 034b9bc8-cef4-4c22-ae27-69e9705552b9 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills Magistral

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:06:19.473510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:a55ad91d57a24dee4f5a2e49cc77efeeabb6561fa0988a37078a4c9ed98ee37d

Observation fe227479-5036-4c4e-9c4a-efa2eac81788 · inbound

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards cites this paper.

Process Reward Models Meet Planning: Generating Precise and Scalable Datasets for Step-Level Rewards Magistral

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-10T04:45:21.193991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T04:40:52.854907Z digest=sha256:3606a657437e03886195d445b81efcb4d891e9d9dbd0d959d264c9efda489e08

Observation 58c87324-1eda-4b9c-9218-514685fecf62 · inbound

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search cites this paper.

What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search Magistral

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:06:05.595526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:19:18.121220Z digest=sha256:638ce3eb74aecd944429c10d02f5a18df6e4799e355958ce9f2ac9e65da01857

Observation 68576317-78a5-4df6-8851-12825ea438f8 · inbound

Language as a Latent Variable for Reasoning Optimization cites this paper.

Language as a Latent Variable for Reasoning Optimization Magistral

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:31:06.613349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T21:40:37.499246Z digest=sha256:86ef7384be815bf0c926ef8f82d3ef21fa29cffb60fc11c13f022c11201c5e3c

Observation 163e5329-0c8b-416f-b23e-05fd51ef83ca · inbound

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction cites this paper.

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction Magistral

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:11:15.966281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T06:34:56.032634Z digest=sha256:cde548a2cb11804b9046838437bf8ab93faeee6b2ea933abd8dba04cc2584449

Observation 3288c138-cde4-4090-805b-18756e46e300 · inbound

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models cites this paper.

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models Magistral

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:06:06.496149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T18:51:59.402906Z digest=sha256:a3f9329a5e339e6a6d06a580e604852ade453bca3e383728b58ae532523a5068

Observation 79c6f8be-3c78-4651-b24a-d1ed5273043c · inbound

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models cites this paper.

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models Magistral

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:55:31.085847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T07:48:28.381924Z digest=sha256:0200b51a5f3438e823b591bd7d5388300376767bba3f082f74b640c9d562cb0d

Observation a26de0a7-cdf0-4b87-9f9f-cda7263cef3d · inbound

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models cites this paper.

When LLMs Stop Following Steps: A Diagnostic Study of Procedural Execution in Language Models Magistral

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T05:20:33.269530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T05:20:33.269530Z digest=sha256:fb5d45a19ea031f553a466287fdb0160c111464d9bf3930d6d714967100b7d83

Observation 8ba3c1fa-e843-4348-b259-15b6efe9335f · inbound

Self-Supervised On-Policy Distillation for Reasoning Language Models cites this paper.

Self-Supervised On-Policy Distillation for Reasoning Language Models Magistral

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:43:21.816674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T14:42:55.368104Z digest=sha256:d444e4bbe3e7d16bef2ccb40e53b20f0ab27caf6c3e07727cdeda06305d40258

Observation 3b6793d2-7bf7-4271-80e8-b6900abc0ea9 · inbound

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning cites this paper.

How Much Thinking is Enough? Quantifying and Understanding Redundancy in LLM Reasoning Magistral

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-05T10:20:57.186398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-05T10:18:18.717871Z digest=sha256:c2a3d7c31ab62e245ffef353af22cb0966c964817c254d54720900cba0b14ad2

Observation 1b858ce7-4e44-4d77-a08b-8ca5bbb14aa8 · inbound

A Primer in Post-Training Reasoning Data: What We Know About How It Works cites this paper.

A Primer in Post-Training Reasoning Data: What We Know About How It Works Magistral

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.508069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T14:40:21.583101Z digest=sha256:fe31490b2d8a662c616d05c83d254edd2a4774d6f6e6c0309d0c7ea516dffa20

Observation a2212d16-e9e4-4e84-9124-6dc20dae2279 · inbound

The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs cites this paper.

The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs Magistral

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:07:13.147843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T22:08:45.988451Z digest=sha256:65b11acf9a2ca0ad1992e6189627d7973506e89fe107bce573f313a53a1673be

Observation 25d944ad-0225-4817-aa40-72d6a6956ee1 · inbound

When Rules Learn: A Self-Evolving Agent for Legal Case Retrieval cites this paper.

When Rules Learn: A Self-Evolving Agent for Legal Case Retrieval Magistral

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:58:47.158738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T03:35:30.841853Z digest=sha256:fd7d6ba754ee574cf3bd6528d8b09218b515279ab094e0e9c36a87327b9a2a45

Observation 3658b2d5-21c7-4ed9-b73f-9483efdbd014 · inbound

Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier cites this paper.

Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier Magistral

Reference 133

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:07:47.307984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T10:36:09.211639Z digest=sha256:c56283714c049034f1e31f3014caff536ddf8e9ad04ef432f4714bae8f9c8ab0

Observation 91a0ab2b-7c07-45aa-9722-061ead9fff09 · inbound

Predictable GRPO: A Closed-Form Model of Training Dynamics cites this paper.

Predictable GRPO: A Closed-Form Model of Training Dynamics Magistral

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:45:29.603748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T06:40:47.911021Z digest=sha256:9f27a766685986ad93946a0f4384353853a219e08d9fee5b3943202b17260ce6

Observation c4f07051-995b-4380-ae9c-81404c659281 · inbound

Predictable GRPO: A Closed-Form Model of Training Dynamics cites this paper.

Predictable GRPO: A Closed-Form Model of Training Dynamics Magistral

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:27:21.687480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T20:26:15.667646Z digest=sha256:d3c6e83b618e303a56be24ba851790991c2dde0fb142d13e2fd9553881fbffad

Observation 7dbec6d2-9bca-4ca5-9c09-a0e8a27e7e55 · inbound

Cost of Reasoning in non-English Languages: A Case Study on Japanese cites this paper.

Cost of Reasoning in non-English Languages: A Case Study on Japanese Magistral

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-14T14:12:44.286699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T14:12:44.286699Z digest=sha256:eb9a18f9ae2316df9870f4b336b945b48a4ac16f8a53a1958f3e422182619f87

Observation 1941e2df-6267-42b0-98bb-c2f57dc50f90 · inbound

Training Large Language Models for Self-Explanation Faithfulness cites this paper.

Training Large Language Models for Self-Explanation Faithfulness Magistral

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-01T08:36:30.534949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:36:30.534949Z digest=sha256:359e5d2e92b183332577ba99ca6f8e1dab4e296fd2fb2c7cb6817f50d94edbe8

Observation b9076e27-ef59-4b11-8172-d2e50da442da · inbound

Contrastive ESA: Human Evaluation of Multiple Translations at Once cites this paper.

Contrastive ESA: Human Evaluation of Multiple Translations at Once Magistral

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T11:48:50.895042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:48:50.895042Z digest=sha256:58004e23280afd93467cbf084a4fb6c0cecb0b6282f7e123ea19bc7f76236925

Observation ebb49bd1-9cb4-4867-8d6d-5f47806d1d42 · inbound

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability cites this paper.

Trustworthy AI in Digital Health: A Comprehensive Review of Robustness and Explainability Magistral

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-04T10:56:20.082102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:56:20.082102Z digest=sha256:8417451afb3af55917b8f25c057d4d502735c6cc07a399cc1cf2253a9180b823

Observation 63f9a6ff-2722-4a98-bdb1-3722d656b668 · inbound

Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning cites this paper.

Beyond Full-Model Rollback: AuroSFT for Adapter-State Multi-Task Fine-Tuning Magistral

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T17:11:01.995426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:11:01.995426Z digest=sha256:258a66b0112b54add9da70fd70d5921c1626686bef5aeaa5b816ccbc51d9a22e