Pith. sign in

Paper Citation Record · LEDGER

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation

As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2507.09850.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.09850 v3

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:51:18.445576Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7633559d-e6e7-49d6-8174-061d73ca5f90 · outbound

This paper cites Contrastive Chain-of-Thought Prompting.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Contrastive Chain-of-Thought Prompting

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:16.980286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:16.980286Z digest=sha256:292b6570b1f2735438fee68446aa7aef9dc3a07b3714d688444270a5fbd3cae2

Observation 6cb8ee24-caf1-4f41-a2c3-d551fbd8ef0c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.164619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.164619Z digest=sha256:c889daec5220db004350b84ca76b6bf82d84fa719af72287fd487161842bd227

Observation f3944f59-10fc-467b-9455-de2c0df293fd · outbound

This paper cites s1: Simple test-time scaling.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation s1: Simple test-time scaling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.567757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.567757Z digest=sha256:0d6669e48d780b33971f1e2b1e8dec07ab2a563202f4f8af951a85461030fe9a

Observation 89346af0-55f5-431c-aa5a-9756990d1cbd · outbound

This paper cites Chain-of-Thought Reasoning Without Prompting.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Chain-of-Thought Reasoning Without Prompting

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.807073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.807073Z digest=sha256:790bb1b278bba34030f3f05a3f2df668115516f0255ccc86d76594ea59689ae6

Observation 7e278407-41c0-471b-9ac3-95814c0ed2ae · outbound

This paper cites Qwen2.5 Technical Report.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Qwen2.5 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.900304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.900304Z digest=sha256:bfa3740db03bc74a7c44b3bb60bbc470292e995b980e1070e7cfa81780d13f0f

Observation 4b3b4f80-3ef6-4235-aed0-c9a318962226 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.971403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.971403Z digest=sha256:f5d4913f9bc231fa4a5f1631f9c21f7ae9f0ccd6b1d3546c2bb8aba15637727f

Observation b3ded519-52c8-419b-8937-bbaaa9dccd8a · outbound

This paper cites SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:18.054670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:18.054670Z digest=sha256:4ba7113171174c68d1200648b1c772317e0f2ba468e575a8545a0218a86c03bb

Observation 5bdb560a-8df6-4ae0-950c-3e6bf4abd00d · outbound

This paper cites Automatic Chain of Thought Prompting in Large Language Models.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Automatic Chain of Thought Prompting in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:18.153821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:18.153821Z digest=sha256:3be5541ea746c878316e9c0d56a1ed67faed1c4e29e95c9915398f32ab6a8a8a

Observation cc35a9d0-f3f1-4a56-9c7b-ef8c73ed6f86 · outbound

This paper cites Data Generation A.1.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Data Generation A.1

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:19.010536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:51:18.351929Z digest=sha256:37900c25759a828cb5a29f898dce7b8a4a363cd02444591ac47b54ec0911e533

Observation 71a31c61-5e24-4cfb-93b8-39259fd68486 · outbound

This paper cites an unresolved cited work.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:51:18.833812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:51:18.445576Z digest=sha256:8734027b170bae57606530bd5f74d9afe3c42f6231c83d0bbd61bc70918aa540

Observation 0d686b90-4d44-4569-aa29-0b00ae6925ff · outbound

This paper cites AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.434256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.434256Z digest=sha256:89309632268794d48a05d05b70671799035ab8600523e086dd6d5bd396c4c38e

Observation 2c81de72-3dc9-445f-bd2f-2f60b4ad8fd7 · outbound

This paper cites 1.4 Million Open-Source Distilled Reasoning Dataset to Empower Large Language Model Training.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation 1.4 Million Open-Source Distilled Reasoning Dataset to Empower Large Language Model Training

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:18.253925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:18.253925Z digest=sha256:5505ceadb968664b420cd6d0783707d43af87e5766c7f28f9af033c9c3f551c6

Observation a93047fa-a45b-4632-8ee3-edd78f5fda81 · outbound

This paper cites Active Prompting with Chain-of-Thought for Large Language Models.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation Active Prompting with Chain-of-Thought for Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.063408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.063408Z digest=sha256:51a955dc02e6aedc7204ce6a8ee8b2bfce0ea97d6e61acf28bdf28ca982e63ee

Observation d6f2d8f8-f5b9-4855-9ecc-2958f85c6327 · outbound

This paper cites NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.701818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.701818Z digest=sha256:093f7bdf0f027e4640ab5c81910318ba3ca0d77c42d317f79420c2ca0402f8ce

Observation afb432f3-93a1-4898-82d6-2917b06861ec · outbound

This paper cites LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!.

The Challenge of Teaching Reasoning to LLMs Without RL or Distillation LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:17.298387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:17.298387Z digest=sha256:59f98842c911f709e516e6ed1f1a97fe2e26e02597de32142991463e3344daf2

Pith citing papers

No inbound Pith citation observations are available.