Pith. sign in

Paper Citation Record · LEDGER

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

As of 18 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 28 inbound Pith citation observations for arXiv:2506.13284.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13284 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:41:36.479613Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:49:04.825028Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 348a1514-6e0a-4921-924d-a45253850050 · outbound

This paper cites OpenCodeReasoning: Advancing Data Distillation for Competitive Coding.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy OpenCodeReasoning: Advancing Data Distillation for Competitive Coding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.663663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.663663Z digest=sha256:a495d3cf552bc00cebff92a8884b568cddeb1edf3440711075dfb76737c75ef4

Observation 251cd77c-ad37-4a0f-9188-74970e050770 · outbound

This paper cites Matharena: Evaluating llms on uncontaminated math competitions, february 2025.URL https://matharena.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Matharena: Evaluating llms on uncontaminated math competitions, february 2025.URL https://matharena

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.297557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:33.722383Z digest=sha256:c86f7d0dd80605e0ff4a19a59aed84d62d53fab5c68f8a7bc0853b53779a54ff

Observation 9854e4b3-d6fa-48ee-b962-b75be4457496 · outbound

This paper cites Llama-Nemotron: Efficient Reasoning Models.arXiv preprint arXiv:2505.00949, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Llama-Nemotron: Efficient Reasoning Models.arXiv preprint arXiv:2505.00949, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.785649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.785649Z digest=sha256:8a87b376139d1bee1bb0cfff370753a3f266724a205c7d3b39abac23eca777cc

Observation bdb8ca73-c91c-4ec5-a086-5b469694843f · outbound

This paper cites Evaluating Large Language Models Trained on Code.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Evaluating Large Language Models Trained on Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.875463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.875463Z digest=sha256:ca3f074ee58822d88e86bf1dee6615d82328aef8ed1cb000d44ed7377aeb1e8a

Observation c2edf7f1-ed2c-4a8f-80a6-07974803d571 · outbound

This paper cites AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.940080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.940080Z digest=sha256:50804d3b26e82a9c136d182f1fcca5b171cccb57ebd149b2eea633aef512eb41

Observation ddb7e23d-d735-4ec3-a22f-3d838b10963b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.004261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.004261Z digest=sha256:36c458758be6694f20dbd8d0a305dcf4b7404271ef69ccb0519760fae4186434

Observation d4944494-ca7c-4a38-8527-d901cc9d6985 · outbound

This paper cites NVLM: Open Frontier-Class Multimodal LLMs.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy NVLM: Open Frontier-Class Multimodal LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.085376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.085376Z digest=sha256:9b4a7a023867c556f0229f979fd70c49104588f512e15e539a91786b28325214

Observation d6eb32c1-0faf-4999-86af-f60a5b54c463 · outbound

This paper cites Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.161298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.161298Z digest=sha256:78f6f1eaf278677f5bfaf11ac075d6098a375898da7b790e0823162449ffc049

Observation 19f5eb4a-102d-4841-ac28-983ecd9d139c · outbound

This paper cites The Llama 3 Herd of Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.233620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.233620Z digest=sha256:db708d8e0a9a077cc90eb6f2eb4fb0e5909bea1499c5f49c78b4a3e11c6e9d57

Observation 11c2f08d-8cfd-4880-b0c5-e25e5580c688 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.293552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.293552Z digest=sha256:0b87c0264b488c2470543403bb44a2232d8eafbbfbc4eeb8af9c123b1dc76475

Observation 9aafee33-eba3-4100-aaed-5c0c03905092 · outbound

This paper cites Skywork open reasoner series, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Skywork open reasoner series, 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.154755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:34.376071Z digest=sha256:b57f5f8aedd93bbbdde9d17bbb2b19d11e1ae9a2cfa1dafcd4d858f8a2645a33

Observation 1ea21ee6-5cfc-455c-84af-e57f4e4a53e8 · outbound

This paper cites Skywork Open Reasoner 1 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Skywork Open Reasoner 1 Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.446923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.446923Z digest=sha256:98072388034e2672e52c5c6ebb23940410561e0877fd0a5a13fb97902128e632

Observation 8ac4abad-be37-47b2-903f-749c58a81629 · outbound

This paper cites Measuring coding challenge competence with apps.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Measuring coding challenge competence with apps

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.048502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:34.541666Z digest=sha256:d85ebf806a1338605803011591ce8c2ad6ebb0674aa8bf59ced36e54fae2b756

Observation 125766b4-f36a-4f80-b7c4-f054e7b61b28 · outbound

This paper cites Measuring mathematical problem solving with the math dataset.Sort, 2(4):0–6, 2021.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Measuring mathematical problem solving with the math dataset.Sort, 2(4):0–6, 2021

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.839167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:34.631933Z digest=sha256:06d83ceeda13634782d2dceaad652bd05a3b0e88925865c2b87b59090c62716f

Observation 002ae11a-2cff-4bd7-98d6-c45ae10de270 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.685646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.685646Z digest=sha256:ce727f1704950627f96ff0c67d5d25b341fc4aa6e298252191e74c19c1a5b4ff

Observation 5d3b1ec4-730b-4bef-b208-bb4a2e515337 · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.657016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:34.781301Z digest=sha256:9220d12c449f66e5b85eed2f98fc4b8379d394198968bbf56af1c1a520afd18c

Observation c4d8b3cc-da20-4cf2-9e67-aad68a30c6f8 · outbound

This paper cites Qwen2.5-Coder Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.858625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.858625Z digest=sha256:b88615bd73028dd07e21196e7b0ed1a0faf933bdc1241f73f19e0bcc368d3953

Observation e0cb5a1d-9dde-429b-b25b-eedf95fd2419 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.924620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.924620Z digest=sha256:3d5c043d4888a5230e05be007ffc24ecf296d1f704aeeda724a839e4f78f75b4

Observation 0f726602-e26b-45fd-964c-7a051a721913 · outbound

This paper cites Numinamath.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Numinamath

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.477517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:34.978623Z digest=sha256:778668c50b79b910c61994a2c0de07efbd57a6f1bee699a7e27ba73ea022de72

Observation 87136dd8-ece5-4de9-ab84-a61df9295c3b · outbound

This paper cites TACO: Topics in Algorithmic COde generation dataset.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy TACO: Topics in Algorithmic COde generation dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.095187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.095187Z digest=sha256:b4970a2539d1f5ef14f43e02f9f2bfe11f0526f9dd3f6dce3123c569460412ae

Observation 2ffce109-9866-43db-92d3-61e73d41ebcd · outbound

This paper cites DeepSeek-V3 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeek-V3 Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.160615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.160615Z digest=sha256:aafa7d5414d95bd42b392501539904429a95a25d3facf6521f3d88012f518171

Observation 71f40547-feea-4652-8c01-086e7bd6adfa · outbound

This paper cites Is your code generated by chatGPT really correct? rigorous evaluation of large language models for code generation.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Is your code generated by chatGPT really correct? rigorous evaluation of large language models for code generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.332590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.244979Z digest=sha256:1e2e91932c2037c9fbc686cbbef40d04b5092251b82b952a31c3bba1014d1793

Observation de862560-3c48-4892-8318-73a99d6e4565 · outbound

This paper cites Evaluating language models for efficient code generation.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Evaluating language models for efficient code generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.117026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.328729Z digest=sha256:ced9f9e108f82a786d3af8cbff382720a76afe5c69435eac11ea1ab08ef9b898

Observation 9945ac94-99ee-4cd5-bc09-518bca2d4311 · outbound

This paper cites ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.394606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.394606Z digest=sha256:cca2f01806fcbb2d959b8f179cd7af1de480df36d43bc4a51407f92cb7009700

Observation d6db3524-6140-4563-bca6-9ab0e975239c · outbound

This paper cites AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.455218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.455218Z digest=sha256:85a1976bb1040d687256dd298d39367dc1335f8ea09ac36ab9a2bbe139110023

Observation 2afc52a7-56e5-438b-8c28-8ba909b50927 · outbound

This paper cites Deepcoder: A fully open-source 14b coder at o3-mini level, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Deepcoder: A fully open-source 14b coder at o3-mini level, 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.888394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.528184Z digest=sha256:9719c43793e81c4819b726830da5706d348d73f1feeda639c7db56b702cc6db3

Observation ea4a2324-f54d-4376-a819-b33b50163490 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.645287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.582507Z digest=sha256:16bd3cf91e6916ca1891060b44d1a2c50c6e73df7f0c0df0b202585fadf31d3b

Observation 2babe2d0-3c1a-4d25-a5a2-1d5a3562de6a · outbound

This paper cites AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.629509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.629509Z digest=sha256:cf496f8d5e318b1db661cd98766f6d7338a6a1df87aa58b6bbad6ec96a32b3a5

Observation b79fcdd3-db65-41e7-8170-e304bc684a96 · outbound

This paper cites s1: Simple test-time scaling.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy s1: Simple test-time scaling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.685233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.685233Z digest=sha256:722097ee4907471095b61103a9628ad70c569c99902feddfbb7dfe93675cc05c

Observation bc47100d-b5b8-42d6-b17f-5d18db604ff4 · outbound

This paper cites Learning to reason with LLMs, 2024.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Learning to reason with LLMs, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.405092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.755514Z digest=sha256:4150ba3f06241210d976bcc24624a71dcbc783d14c89d567e780bac259bbb2aa

Observation 1ab9a39b-735e-45a5-8a5c-45d69449b868 · outbound

This paper cites QwQ-32B: Embracing the Power of Reinforcement Learning, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy QwQ-32B: Embracing the Power of Reinforcement Learning, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.162217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.797155Z digest=sha256:99dce51d9b3a8c6f0182f6178068b15c8416755a9483f5021e85ec9431fae5ff

Observation c05608b2-c707-47d4-b036-421817536693 · outbound

This paper cites Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:37.812618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:35.854088Z digest=sha256:f35034e6a410d4dd21bfb4905ea196de9e36884ab6105a914872965091410678

Observation 354f8ca7-2d52-4513-8a7d-67db3398f439 · outbound

This paper cites Seed-thinking-v1.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Seed-thinking-v1

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.912258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.912258Z digest=sha256:f6e4966645495a8546a9bac932960c464452fb713b0dfd2eb34db0a09cbf0544

Observation 2a1a5481-17f1-4988-bfad-ec0c516bf108 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.965118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.965118Z digest=sha256:88cb8f7d9e5159f57099901e61911aee68e1ba6c941fb07002113423c902c5f9

Observation f7566cfb-ade5-4ea6-b269-0b0c8818a398 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy HybridFlow: A Flexible and Efficient RLHF Framework

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.026976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.026976Z digest=sha256:a7ebedcb7346ca19fdd59261db3aba8d76fdb3ffd43c996657030eeafbb06fb6

Observation 93e0b815-2d12-4c8e-a96b-0ec5d8764436 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:37.222993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T00:41:36.056695Z digest=sha256:a3679264f9016f837b74ef52d338e98e3861a4877282f18876ead69fd0b4daae

Observation 71a6e346-aac4-4dea-a418-2222643eb1ae · outbound

This paper cites Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.113617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.113617Z digest=sha256:449fc7aab82d029e8a5cb6ec3760926bd5ab7d544de8405e77d6332adbb96009

Observation 56c7b4e5-03ba-44f0-8652-467e2c577f9e · outbound

This paper cites MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.175903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.175903Z digest=sha256:d9b390e8064622ae6a2cd905c6db08c89d47351dcc4f6dabcb6622defca6c375

Observation 089a198a-d9b2-4258-9c7f-c451936bb0c1 · outbound

This paper cites Qwen2.5 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.237558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.237558Z digest=sha256:b0d35091b8f9c02bc4196b26b6306628b98589452f4f0a124303a68cde8523bb

Observation a02be7d8-538d-4b6c-8f07-a8abbbed94cd · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.284404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.284404Z digest=sha256:a785e6c07e4db9a6ec0f9d66ac81136b74c8b95a48505435cd1798b5bc782064

Observation 653ecf0f-07fd-44be-b8f6-01027a8fe71f · outbound

This paper cites Qwen3 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen3 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.361667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.361667Z digest=sha256:aa32d65117fed4292dd7f9ff302e8878ec3b1c3f56a753798f98221631d2b74f

Observation b93c90ad-17e2-412f-aad4-b3c5f5b4155f · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.411140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.411140Z digest=sha256:0d46809da20267dfd28dae67620ab4e349e7b5638f8cd5a64e663a634340902d

Observation 39ae18ee-d0a8-474b-9222-8cddd87ad256 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.479613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.479613Z digest=sha256:e989c7dd37cf8a372fb8d4ceb0770b64641c0e5738c73931e2deca01fb80b74c

Pith citing papers

Observation e7e64ebd-a01c-41a6-b536-78b606eefee0 · inbound

DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation cites this paper.

DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:52:33.396332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:52:33.396332Z digest=sha256:9bf7dee8a3619cbc954f315d02a2e9e5328b608813f311f128c37a635780f384

Observation 715a276c-f616-4c4e-a226-32d82389387f · inbound

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs cites this paper.

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:12.895530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:12.895530Z digest=sha256:2010a4b58c9ccc21b07b013e6b07913458df6497ebb1ddd29e7ab02e2d692d1e

Observation 263d6ffd-74e9-4133-b666-486e59853516 · inbound

The Signal is in the Steps: Local Scoring for Reasoning Data Selection cites this paper.

The Signal is in the Steps: Local Scoring for Reasoning Data Selection AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:16:14.215851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T10:14:27.739531Z digest=sha256:3d7cc107e51616cbc813a047d0f59e4dc01baaa24e059cba5dc9ff4bd905cef7

Observation ac544d11-8b86-4391-9fb5-28d1cf1c98e2 · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:06:13.902007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:8f4b88625024b3151ec3695f84464000352a26d854d0aaefa9e902ca9b6b57ce

Observation eead6158-8dfa-463a-a142-fcfe2b778c16 · inbound

Video Reasoning without Training cites this paper.

Video Reasoning without Training AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:08.416668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:08.416668Z digest=sha256:e6578235d4802c2e3e8bd8e222b3ff4801b0430d8f2340ad2cc6e94b6cef47ea

Observation afb90df5-e7b0-44f9-8e70-6f30577c4aae · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:43:11.941650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T07:43:11.620446Z digest=sha256:e767b1c8f45711290bb1be0f0e8ea32a77ada780bb7f837da44b1b50fca7458f

Observation d91f228b-8b9d-4d6c-8007-241b3ebc6d38 · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T07:31:49.069930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:31:49.069930Z digest=sha256:37e8a0694760e9d0476f6de1aa23707225b41c0c150b52a66f3d73b133066443

Observation 5d94fa74-f831-48c4-ab64-c5f8bb37ee83 · inbound

NVIDIA Nemotron 3: Efficient and Open Intelligence cites this paper.

NVIDIA Nemotron 3: Efficient and Open Intelligence AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 170

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T01:40:42.712572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T01:40:42.190369Z digest=sha256:599c1e52930861b79449fb4845406cf52415424c3b5aa0410de35296206e8706

Observation f0c6be84-0d9d-4402-ab99-02590d6f2afe · inbound

On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training cites this paper.

On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:48:00.530859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T14:47:52.094353Z digest=sha256:5161746b6530a34b0ba8dab06c6dbaf76c56ac4d1f4572c88b5ee084b9b27435

Observation 386d4239-5b69-4d4b-8170-b11d040d9dcf · inbound

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse cites this paper.

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:47:37.216769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T08:47:29.236561Z digest=sha256:ff1efdd093c794a1fe248de1698470601c9c2e6945727863804e001663e4315f

Observation 1bfe6c99-0e3a-400c-b390-91593e889ac1 · inbound

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse cites this paper.

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T05:50:24.502645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:50:24.502645Z digest=sha256:371811d97d21b21f9ccb796f9a730a52a9b02a81e141a1cf3a63f779352ea72d

Observation 1af9d3c6-ab40-4d68-a453-aac7e9427bf4 · inbound

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models cites this paper.

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T05:28:13.898465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:28:13.898465Z digest=sha256:087fdc3d585fb1bf663d76925f935155c31eb3fa9fb6307790b11e20f1473678

Observation 68335bce-2ae7-42b0-9ef4-251cf552bce8 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.537608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:a3027258afce5858321bf840b4cd0bf996bd0171314c77be924868de4a65b245

Observation 538c34aa-db4b-4a15-a4ed-22f28988bc3c · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.084430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:ba63525fdfa3abbf9bc4cf0513e35aca8c97658829bd6b96e817ce391e7b1016

Observation 6dabc6dd-63a6-434f-ab24-e836b93fdf50 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:06.791321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:06.791321Z digest=sha256:323b8b5b19b5fa14aec018b3758f0ea46a61737f2930abea5c83adea6a4ba015

Observation 9ab6c976-a761-443f-93a7-67ca89d3f58b · inbound

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data cites this paper.

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:18:21.791845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T00:18:03.185565Z digest=sha256:0143126745cdd949f32ecc89106e6c3a8aa87c62b25061cdbdfafb4e754ad1aa

Observation 84f59440-29af-4df6-917f-caa49c812b68 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:56:11.635008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:f70248600bac67a85939eadb4600ab239a4c6a607b620269dbf719cd4478e58b

Observation db8bbd0e-e3bf-45ce-af2f-bb7dc3e87169 · inbound

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models cites this paper.

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:45:57.822136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T02:44:36.697450Z digest=sha256:781b3d29da0509bc2c7eb7c7d34dde3f6d3afa7cfdb80a24f20a1479957a3bce

Observation f0020f30-01b6-496c-9681-05b896a5c9e5 · inbound

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models cites this paper.

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:03:50.645752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T23:01:17.957634Z digest=sha256:00add1e2dd7a938590016ac92cc1d5241ba829910eac66b1d699c9eeef973f6f

Observation aa0f8adc-f8b6-47f9-84d1-1f9495c39b3a · inbound

Post-Trained MoE Can Skip Half Experts via Self-Distillation cites this paper.

Post-Trained MoE Can Skip Half Experts via Self-Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:03:15.213905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T12:00:35.496822Z digest=sha256:a4955240769c64ea752c71281a02df170edc4632684012bea0416d90ef3dee2f

Observation 184f504f-b847-4fee-8101-a1d82d72c29e · inbound

Post-Trained MoE Can Skip Half Experts via Self-Distillation cites this paper.

Post-Trained MoE Can Skip Half Experts via Self-Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:25:00.079365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T18:22:45.702572Z digest=sha256:e0d908ee514b4cbd1c01c5c0cf8e587d58723036d54b99c166624a00fcb4a2a1

Observation 8630763e-b68f-41cb-944c-fdcacb880604 · inbound

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models cites this paper.

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.530928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T08:01:39.412431Z digest=sha256:714adf601c857acd9b29eb9eda21049d456e55d6189e053433152ffb6ff49db2

Observation e1dd7765-87b1-4ff2-9fbb-f236cda1a1cc · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 255

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:56:13.643184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:660a04abd1009d8be7c82adf84b626680b9d70427b072b19d2885ee1c50af4bd

Observation 5941e877-a830-468f-8e55-20c5f836ebe2 · inbound

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning cites this paper.

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:47:59.921475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T10:18:54.163862Z digest=sha256:05543dc3cb3d05843b8a4e04b7819297e55874294248ce124fad66d9faad18f6

Observation 394020ad-4f17-45e5-b4a5-970d706b09fe · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-13T03:06:27.991558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:06:27.991558Z digest=sha256:ef42a654987af5eb0739e599fd912dacf4deb118590ba88f8e4d860be27daa01

Observation 9e440f94-c3f8-4a69-ad08-a6249f81933b · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T15:13:39.458378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:13:39.458378Z digest=sha256:40193bc69719a78c3452049111b2461eb81a90ddf8231cf0bdc70d52b286def1

Observation 081a4e37-aac3-49e0-9ac7-af96fbcc5f6d · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T07:45:34.584748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:45:34.584748Z digest=sha256:b3f5b87de3db7693b9e9d9453b18eb1ca3d6d68695cb899b4527b0c11cec12d9

Observation 5cae5c97-923b-43f6-9f73-e2cfc7521f86 · inbound

Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility cites this paper.

Test-Time Scaling in Reasoning LLMs: Inference Regimes, Evaluation, and Reproducibility AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-15T14:49:04.825028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:49:04.825028Z digest=sha256:aa0537290608cced8d11eea8a5d801a7c9ca13682bb6d0e241956be9171f9276