Pith. sign in

Paper Citation Record · LEDGER

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

As of 8 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 5 inbound Pith citation observations for arXiv:2502.06773.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06773 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:25:53.635679Z

measured 92 of 92 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:41.579717Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:57:41.592060Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved66
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ecc36f18-1e9b-40fe-93ff-1e108ea1af3a · outbound

This paper cites Phi-4 Technical Report.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.261354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.261354Z digest=sha256:48258620a851d1c5be08e011d9675ce9dee80ad2934adb751e47da34bee6c8dd

Observation b582b0e9-a95e-4d3e-9f94-0fb09582e460 · outbound

This paper cites Amc 2023.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Amc 2023

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.267178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.267178Z digest=sha256:b9a31cb88cbbba3f9db13a57c1befdeaa8a7f63a9a26bf226a39d7009f86bb4c

Observation eece90db-4388-4ddf-9c31-7642bde6225a · outbound

This paper cites Aime 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Aime 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.271469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.271469Z digest=sha256:7ee359712d20277b5efd1b3cb1d19e2695f370b6078164c22f02aacfebdd5362

Observation b5f48994-4ec5-4212-9c7f-2cefbe971e23 · outbound

This paper cites Numinamath-cot.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Numinamath-cot

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.275821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.275821Z digest=sha256:a6eb020d713a8010b447c526f71848b88ab3243fbea7d4da1ba2ed7c60fcb18e

Observation 6a10645d-8bfd-4e75-b884-8f82f92de4b7 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.280245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.280245Z digest=sha256:e02905eac90cd1178880355b6841cff996da3eab6fd85f19bfda60f3e7451855

Observation c281d1d5-4ccc-459e-985e-eb611e13f7f9 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Constitutional AI: Harmlessness from AI Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.284895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.284895Z digest=sha256:f29e6da6762aa4eb07ffe029625e5510ce63ec91b9cb3bd3e7e84a58dc19957b

Observation 1082ab03-ef0f-4ca1-97a8-9f0e7f0cbfcb · outbound

This paper cites Scaling test-time compute with open models, 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling test-time compute with open models, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.290041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.290041Z digest=sha256:c9585fb26e846b454cc986cc277629805dba64142ff486761cc880b8e3376938

Observation 0b217de4-d4a0-4bb4-ad06-04e7fdc930c6 · outbound

This paper cites Open-r1: a fully open reproduction of deepseek-r1.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Open-r1: a fully open reproduction of deepseek-r1

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.294520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.294520Z digest=sha256:426a2fbcee1c1eae583bdfadebf24e6ca90d669a2dccd0fa848f734c79dfeeba

Observation 82bdc4bf-793f-4cd1-ad2b-fd0c52762730 · outbound

This paper cites LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.299016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.299016Z digest=sha256:70163667f9cfb7ee8426331c9e2128192689367f72cffb4a58b4f2a5a53c8fc9

Observation 3357ea57-02a9-44c2-937e-6a19461ad863 · outbound

This paper cites Self-Improving Robust Preference Optimization.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Self-Improving Robust Preference Optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.303921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.303921Z digest=sha256:6127fe640f48d12c208e367ba25d7af61c7fda34dcacf8c31c82a10686286347

Observation 2402dfe1-17ca-45d9-8aa6-e55610864943 · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition AlphaMath Almost Zero: Process Supervision without Process

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.308231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.308231Z digest=sha256:003c689d20aec2699884f371274982f7707ad20ccde46f4f8d01998bc5ea7af6

Observation ce1aee1b-0de3-4e4b-9af5-2daeb8862248 · outbound

This paper cites Codeforces dataset.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Codeforces dataset

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.312356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.312356Z digest=sha256:7b38c5211df2d027194db85853f8e806c115519729b2b461311a275e71a1f8fa

Observation 1e9f82f3-63a8-4862-bab1-3583f27d8217 · outbound

This paper cites Process reinforcement through implicit rewards.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Process reinforcement through implicit rewards

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.316251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.316251Z digest=sha256:25d65dadc36ca1f4435e2031ee959850d17de56d032c0191f5ba95a927bb3f34

Observation b2b9e052-b993-479d-b1e8-f235fbb6df7b · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.891480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.320253Z digest=sha256:ea6715bbe03bfbbddabd846f706c93c8fb20218bda1566ba9f25aa227a4f3a6a

Observation 51617f6d-02d7-4ade-b1ae-f7ee7d57f972 · outbound

This paper cites Flash A ttention-2: Faster attention with better parallelism and work partitioning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Flash A ttention-2: Faster attention with better parallelism and work partitioning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.875355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.324290Z digest=sha256:d30e9efce43ae3fabe8f94133ddb28c18329f2bd737651b0b7403b121d43e59f

Observation c4041b13-e1a1-492a-9fae-53fd468c599b · outbound

This paper cites Ai achieves silver-medal standard solving international mathematical olympiad problems, 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Ai achieves silver-medal standard solving international mathematical olympiad problems, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.860648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.328130Z digest=sha256:76e428dfd1684a994d4f5dc183d64f572940f21fa8255ab164a96e11061e5aee

Observation bdd185cc-189c-4585-a351-a8aeff4a8531 · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.332278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.332278Z digest=sha256:3c6ccaa65e990c9e35eca5c0df70c54033bd9eb13f0c8a033484c50fb4bbc026

Observation 56005d59-b58d-4a7d-a27f-8341ea7a820f · outbound

This paper cites Introducing gemini 2.0: our new ai model for the agentic era.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Introducing gemini 2.0: our new ai model for the agentic era

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.846667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.336397Z digest=sha256:7bcc31f8b3c812293c201d01cf603ea255e927a2b5f7811c4feb7e116d0e81d7

Observation c7e1636c-8b03-4263-91b1-ae5239e0bc6a · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.340082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.340082Z digest=sha256:6b9dee4eed81a9369ccb4e1f755b668f7968d928dae67e8a825f8cdab4fff881

Observation 6d640c67-b92c-4e0d-9e83-c349fecca3a1 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Measuring Massive Multitask Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.344239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.344239Z digest=sha256:dbd27410e3f0021232134eb13b6b881aee57dcea7e32e0507ff880d0f1b36a93

Observation 388a53cd-7a7e-43c0-90f1-bcc3a536fd76 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Measuring Mathematical Problem Solving With the MATH Dataset

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.348541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.348541Z digest=sha256:4a1f919d6d1b518904a3ffeb0b2e38fcc2f9877cddb1506214a5c0303ad6a15b

Observation c2702d0c-f576-4c21-8bb6-35cca6a3d8e0 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Large Language Models Cannot Self-Correct Reasoning Yet

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.352790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.352790Z digest=sha256:1cfb7d9f21d675788edc828bca4062970179917dd9d2f552f3f5fd9a7fe7c944

Observation 943e4545-6612-4952-a7c6-53c23682a60a · outbound

This paper cites Teaching Large Language Models to Reason with Reinforcement Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Teaching Large Language Models to Reason with Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.357179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.357179Z digest=sha256:767686cf3b3a595e5dafcff6346a526470616833d607eff073894318122cc39a

Observation a28b3f64-5a69-4c93-831e-8e7869e8ff38 · outbound

This paper cites O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.361801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.361801Z digest=sha256:3c85a3323d2f80dc935ed9716c2d315b1fa54298ae4f5146b4c5f78059351ee0

Observation b1e892ab-f379-458d-9888-707cb8c2c758 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Reasoning with Language Model is Planning with World Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.366490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.366490Z digest=sha256:46acde99fc8f30201c75373a91fe3459553e5badde83e67f63203b1060b7d447

Observation 5c92a674-1136-4135-92ae-a8c9385107c5 · outbound

This paper cites GPT-4o System Card.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.370738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.370738Z digest=sha256:ac1732161c3d80d9196d50be4c4398040f1aabfde846cd6622edbd82ff683b06

Observation 89738edf-e17b-4fb8-aa32-0e92980a4dfa · outbound

This paper cites T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.374934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.374934Z digest=sha256:9a181e7bd3d73daafaa75be041fd033030126e2194844beb4366ab8612826aee

Observation 71c91751-5438-45e2-bc14-697b74db75eb · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.379420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.379420Z digest=sha256:676d41969a078d5fbf775bb27372675bb5e01c99809568da423afea3348f05a5

Observation 6b593b83-c156-43af-984f-0629a2a6616e · outbound

This paper cites O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.383766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.383766Z digest=sha256:8f2ccb9fee027c58a945b8a7910010de330076e1301cf3e67803018dd11f399a

Observation 9a181384-00a6-44af-a6d7-0d6ebdc467e3 · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.387950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.387950Z digest=sha256:4cd5860c804b5ff2aa43d1364adeb5f3e979188bab5a3b31bb2de4dcbad9e936

Observation 03b4dae8-94e7-44d4-935e-32a759b793f1 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.392275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.392275Z digest=sha256:09ebdd9dc401c3ffe1225f483cfb25a73096120c38658538cb5d6898d91244b4

Observation a9ce09a3-0a28-4bbe-baa8-a3c690682860 · outbound

This paper cites OpenAI o1 System Card.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenAI o1 System Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.396401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.396401Z digest=sha256:5917e056cff3f461d87d9e922cc048bd7a1c068a42b830c36325ac1ba0f6c48a

Observation 9fd80777-38eb-4813-b345-735b42bb596c · outbound

This paper cites Reinforcement Learning with Unsupervised Auxiliary Tasks.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Reinforcement Learning with Unsupervised Auxiliary Tasks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.401082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.401082Z digest=sha256:19bdfffc1d71b2e6621dc935fd6d9626cb3d1da46a3e5f7de52ed7dc02a25229

Observation 9101808f-b24d-40c7-8fe0-e762b331cb81 · outbound

This paper cites Kimi k1.5: Scaling reinforcement learning with llms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Kimi k1.5: Scaling reinforcement learning with llms

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.832565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.405463Z digest=sha256:6a2fc7c0d099df155e72d4ca1455a1d330043fad7c303ee6de27523876e05a2b

Observation 79363580-8c3b-4201-904c-95637ea29156 · outbound

This paper cites MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.409568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.409568Z digest=sha256:0a60e2ad9c1282ab051a7da986503f48c1d1d8825ac5b41a76d37e726c3f161e

Observation ab2f45c7-cd1e-4198-985a-51258f2f78d0 · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Training Language Models to Self-Correct via Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.413796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.413796Z digest=sha256:a776776bd5e1c1b43da54bf436e6840684b34cf703d7356b8e3f493856da9714

Observation 30c427f1-2c02-460d-8113-97c8523d761f · outbound

This paper cites Numinamath: The largest public dataset in ai4maths with 860k pairs of competition math problems and solutions.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Numinamath: The largest public dataset in ai4maths with 860k pairs of competition math problems and solutions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.818246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.418178Z digest=sha256:66e9be6ff809b2b2035ccd0dcb8873e6d48c5168037b347f777608e4c92cc74e

Observation dd0171c7-9298-42b6-be98-39bf2838d347 · outbound

This paper cites Let's Verify Step by Step.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Let's Verify Step by Step

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.422225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.422225Z digest=sha256:4831591e41a77f25997826901591cde3ac49020ac294e3d2de3c65e7fd3a198e

Observation a58e83aa-0365-45a6-87bb-477e0bcc5451 · outbound

This paper cites Chain of Thought Empowers Transformers to Solve Inherently Serial Problems.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Chain of Thought Empowers Transformers to Solve Inherently Serial Problems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.426639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.426639Z digest=sha256:bdd3599651101b4da3cbfcfcd9b0ff08f50aa237a093fd6fa6f184b064411afd

Observation fe092817-e977-4101-87d0-9c9ab5f6a8fd · outbound

This paper cites WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.430999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.430999Z digest=sha256:330e3ffd31bef74542c1f2efcee00c07f93282d1f86abf9874793df6afea4cb7

Observation 85937927-b0c4-406b-bc68-afea3a066a6a · outbound

This paper cites Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.435670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.435670Z digest=sha256:8e342ef38279064b9e28d6ae0d941b99957da725a3a81ebab7d8860af80246fb

Observation 6ad11a03-cdce-412c-842b-a4a40d04aaa1 · outbound

This paper cites Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.439518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.439518Z digest=sha256:7f755f008ee65938c4965a60ab052db4b5061f7718f179463a43dbe08291c1da

Observation 15b8f081-e7f6-412a-8cb6-8426317c24a6 · outbound

This paper cites Llama-3.1-8b.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Llama-3.1-8b

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.793295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.443750Z digest=sha256:c6a139ef46fecb4546e48ac7911792e8f4cb7215b069f57be3afca8d58bc4a6c

Observation b208dd9e-f17c-4bc1-a706-1b09b8c00692 · outbound

This paper cites Ray: A distributed framework for emerging \ AI \ applications.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Ray: A distributed framework for emerging \ AI \ applications

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.779383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.447529Z digest=sha256:3e5f690d0be5cef152bd3f136fb65ccf6b77f06eea279b06dacfae8c44c0962b

Observation ac323491-468e-4ade-8b7c-7ae2a9fea4f6 · outbound

This paper cites The Expressive Power of Transformers with Chain of Thought.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition The Expressive Power of Transformers with Chain of Thought

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.451451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.451451Z digest=sha256:60c3869cf381f55f49189c0dc8d809505d6c816665245171e2a8036b8de8c926

Observation b2b5d512-62b9-456e-893a-26c3ce31dee5 · outbound

This paper cites s1: Simple test-time scaling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition s1: Simple test-time scaling

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.455635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.455635Z digest=sha256:8ed85b5ebbcfd9c82c2909ea37f8c901f157c8cbfcf10ef6411cc4cac3212b41

Observation e3e9fde9-2d69-418b-9f3b-12c5c3db937a · outbound

This paper cites Sky-t1: Train your own o1 preview model within \ 450.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Sky-t1: Train your own o1 preview model within \ 450

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.764289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.459617Z digest=sha256:d03b97c755275d22e32e3dd57561c1007a29af450a5a966add29e310a464e30a

Observation b7089b6b-670b-4e80-a58d-dc265ff58890 · outbound

This paper cites an unresolved cited work.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:25:54.748210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.463706Z digest=sha256:b701c7eb10674a49f455c41e070b8092595b32004a6bf6e95bcbb2d34e4fa250

Observation 72bc56d0-1ef5-4f09-ac0f-67b5db078475 · outbound

This paper cites Learning to reason with llms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Learning to reason with llms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.733808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.468071Z digest=sha256:fcf69cc3df28823477d59d2335b17f7a170cae4b199d4f4fef064b5b96080ba3

Observation a0a00aeb-efcb-44cc-bc5b-b1de23dd640a · outbound

This paper cites Math-500.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Math-500

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.719735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.472120Z digest=sha256:d1fb6aac0e647f034cfdd2c68ed0680c426e6fed7bceb29a277b5f4b2a88a000

Observation dce3a3b1-830d-44f5-a017-d2b828112f30 · outbound

This paper cites Openai humaneval.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Openai humaneval

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.705136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.478200Z digest=sha256:4c8f6792f0c5ef8426d25ab79eed462d635e504c5c69c4a412819257d7d51aac

Observation ecc12a85-1bea-422a-a208-8604b0edfa99 · outbound

This paper cites Openai o1-mini advancing cost-efficient reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Openai o1-mini advancing cost-efficient reasoning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.691110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.482502Z digest=sha256:bd20dba57f06d92e752079fcc654b76fc9e038d5ef9baddc779b89d882c846f0

Observation eb58fd79-d437-4c38-b627-6b5ec21ecaad · outbound

This paper cites Training language models to follow instructions with human feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Training language models to follow instructions with human feedback

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.486672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.486672Z digest=sha256:94f3b002c2e44266e3f5a2fa691d207c05aae78a67e57d6c9617f074276a822e

Observation aa5ce6f6-2465-4fbc-9d8a-af09777caf75 · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.490904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.490904Z digest=sha256:67c86dce5efd852439942919c0715844227dd2e7da5158ef1b5979fb1feb86aa

Observation 8b4943d0-f8c2-4144-99f0-d7676e972f86 · outbound

This paper cites Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.496098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.496098Z digest=sha256:1407a42a946bc1b03b2d165273f60a552e9412456322970db438013f669647b9

Observation c1a0fb65-ef3a-426b-8805-2aef8b147292 · outbound

This paper cites Qwen-2.5-32b.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwen-2.5-32b

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.665820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.500841Z digest=sha256:8f0e16314f5422002029302ca0c1faec2b0e55d328a4824c8a2a00cd688d59ef

Observation 1283d6a6-60a1-46c0-ad8d-ccc8592df38b · outbound

This paper cites Qwq-32b-preview.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq-32b-preview

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.651382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.506146Z digest=sha256:be9e28a46f857dc25bde6af9041fe47fe78481fd2bd36035ddd38ba7363948fc

Observation 8ade35d2-357e-4fcc-906c-02673a1a5888 · outbound

This paper cites Qwq-longcot-130k-cleaned.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq-longcot-130k-cleaned

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.635311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.510751Z digest=sha256:427ad404c60b4d2580262d6aa8977deac89fca932850460aac248bc6112d5f7a

Observation b67dcc04-5003-4918-a1a2-cff852db7ea0 · outbound

This paper cites Qwq: Reflect deeply on the boundaries of the unknown.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq: Reflect deeply on the boundaries of the unknown

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.620823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.515039Z digest=sha256:4b5a1948168489b9f83d9c8f8cf989abf3d6ba115b4c8beede73751928d85567

Observation 41f5cea6-2626-484a-90ac-1707ef95cbc1 · outbound

This paper cites Recursive Introspection: Teaching Language Model Agents How to Self-Improve.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Recursive Introspection: Teaching Language Model Agents How to Self-Improve

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.519496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.519496Z digest=sha256:2e0db4ffcbb1c9c33c62bb3d51280c83a8ff73759326688f0ea17fc25da73ce1

Observation 73ac72b7-cd88-409a-8e9e-983c7e3826ee · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.524212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.524212Z digest=sha256:111d330755bd4f570434127bff1eae69b70403e15d626501eda864527cb0dbdd

Observation 118b5f30-a0b3-47a3-8ce1-de26512b8c3e · outbound

This paper cites A program for the machine translation of natural languages.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition A program for the machine translation of natural languages

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.606496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.528679Z digest=sha256:c3be18d36b45f81e5a0603607a9e2b97f4e81ba1c89aee40b6005a1bc2ebcda3

Observation 1a275c71-3f7d-4a64-b783-e466d9d6137f · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.532654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.532654Z digest=sha256:241da4d357c49d1a6dea4e685d21de3e64ad764726f995cd15bcf5295d3e137b

Observation 46209a82-f91d-419f-94a7-731aeaee85d3 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.536838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.536838Z digest=sha256:138c98a123fcb507e66502311fd2a8940941d3a1fcbc61833a3dfbf04bbe42e9

Observation b8806b55-4f4f-45fc-819c-638cf7f0127a · outbound

This paper cites Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.541676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.541676Z digest=sha256:28fbef97698a324d41b652c339a639a38c52eda5cd5f08b951fe1efcfafbeddf

Observation 568c3934-2dd7-4746-81f8-88f9b644773d · outbound

This paper cites Proximal Policy Optimization Algorithms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Proximal Policy Optimization Algorithms

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.545928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.545928Z digest=sha256:1c2e871ffbfb17139f38967bcec2171e3631b1609f32e94079ed3ddbc445afaa

Observation 400e6bbb-f577-4f4a-91de-f675b930a8bf · outbound

This paper cites Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.550735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.550735Z digest=sha256:7d802d5f48122d3ea7008fd65ba513f05f9f1f2b3a3cfb1719412af810c7cc82

Observation 26cbd230-74a1-45e0-90b6-77d1584b1b29 · outbound

This paper cites Solving olympiad geometry without human demonstrations.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Solving olympiad geometry without human demonstrations

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.591672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.554997Z digest=sha256:91d4ba4aad39027a5d4c05035b7a6a03f85b9a20aa8426dc14374707a08b14bb

Observation c967d194-9948-4029-b6f7-2b1915b48a62 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Solving math word problems with process- and outcome-based feedback

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.558980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.558980Z digest=sha256:9111168e6e0da363c2cff860f17babdd6f2cc504f152611b43f456a479fdd306

Observation 078bb60e-14d7-4bdd-8e97-d4c805a86e85 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.563368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.563368Z digest=sha256:04ce71fc51dec6d83bda2303b37f5b913c55c0b8db80052c6bce4801cc83e9b5

Observation 02e6b3c6-4fce-4229-bd2a-91e10e91dc0b · outbound

This paper cites Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.567513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.567513Z digest=sha256:1ff018160a0231b9bffba75f37f858ec2ef62fe15f9fa136d064dcff9d619e91

Observation 8889cdf4-a380-4569-9c9b-92db64076048 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.571619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.571619Z digest=sha256:074e818d29d3a72208a39fc32db408b84ad32ad200f3131eeb054adf3d03c568

Observation 8c7e0202-04a5-4727-a56e-85f8bf1c08e8 · outbound

This paper cites An empirical analysis of compute-optimal inference for problem-solving with language models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition An empirical analysis of compute-optimal inference for problem-solving with language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.576399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.576225Z digest=sha256:06be812c73c20b342a9595c604a6a50da3127e8451b498f8d4a9c6c44bbd7c55

Observation 490f703c-c4cf-40e5-a257-a31153a9bfa5 · outbound

This paper cites Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.580233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.580233Z digest=sha256:c42e3a743b882b436d103349b66f224de6777f4b2bb2ccf42de3a0874381b4f8

Observation dc1766c5-126c-43cc-9c11-535b88589061 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.584409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.584409Z digest=sha256:a4da0f02657d3d2e5210f96ea682368904ccae81856e3a51beda475e63dedee7

Observation 1fae56f8-5d36-4880-b537-d2c5179ab3c5 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.588274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.588274Z digest=sha256:8cca89f7a34e96f44680f0e769eecd815a802c4d149353564c96cea179f10901

Observation 04a629f7-ff2c-4e73-a168-ca0f29f98e55 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.592132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.592132Z digest=sha256:784239e9e15ac3a2fa3a2e9491db88006e7f1c5eb01c636398b7c53e7b6f9590

Observation 62e15922-f60b-4ea6-811f-e6b923915a3d · outbound

This paper cites MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.596276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.596276Z digest=sha256:55bbecf0633c8f928709a7bc06765fca64f99d7bd86f4806162ca1853245a74e

Observation 9754a7d2-d866-4e3e-9803-e7d1ff172adf · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.600448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.600448Z digest=sha256:95e8c98939409eb67454825151dfa3d0a4743d8907a30c9dfaa39d6f0a6e5634

Observation 6a28b955-a2f4-4089-9f84-33f32dc4962a · outbound

This paper cites Scaling Relationship on Learning Mathematical Reasoning with Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.604965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.604965Z digest=sha256:d6c5ce9e3f41e1d0607a135eafc30fb949527edd0cd4bbaf08086ae0651ce4de

Observation 47af6924-d35a-47f8-b9c0-07fe3f934934 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Tree of thoughts: Deliberate problem solving with large language models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.609543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.609543Z digest=sha256:6893dc7a8ba3a09d05bfa6ce093acd01276cbee3090cace4d33fd303d39cd1af

Observation 1049b08b-fa90-4bb7-b99c-ebca1098589a · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.613796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.613796Z digest=sha256:7ae553bc091d8ebe10db5ab35614cc9b91235d6580fbed95b5f8d6ad5e5ddd30

Observation a6936156-37c3-4b86-a63d-0d780164f52d · outbound

This paper cites 7b model and 8k examples: Emerging reasoning with reinforcement learning is both effective and efficient.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition 7b model and 8k examples: Emerging reasoning with reinforcement learning is both effective and efficient

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.549628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.618399Z digest=sha256:59a59f80454e987b0cb964fb539f781374298ed2604dbb64b31ff9aea63c21d0

Observation 0f9bad28-0108-4b8f-bf7c-123fe2271a93 · outbound

This paper cites Small Language Models Need Strong Verifiers to Self-Correct Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Small Language Models Need Strong Verifiers to Self-Correct Reasoning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.622557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.622557Z digest=sha256:cf324867327fc54be742552f98afc3e62348db3ac01e8f38f6b4ca3fb9dc3083

Observation 388520e0-f5bf-489f-a245-0ef8e55cbf75 · outbound

This paper cites Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.627005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.627005Z digest=sha256:b6a65e1d7db13813b3fb300c3f474902c9a821445e68cd141b5d1d24a96b62f3

Observation 407f26e5-679b-458a-860b-e946dfad7d99 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.631264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.631264Z digest=sha256:f8daff27faf5620ae0a1460b65134de565731b76658dafcfff49dbd501df2e5c

Observation e4199f7f-f4b0-424c-8252-8677147d40fd · outbound

This paper cites ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.635679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.635679Z digest=sha256:a92285bcd3a53e5a3a3d1cf8b7609ee9a6f050c0b4d65d769feef1c9a42ddcf8

Pith citing papers

Observation 9c08c4c8-1eff-47cd-8605-016f003dce04 · inbound

Phi-4-reasoning Technical Report cites this paper.

Phi-4-reasoning Technical Report On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:40:25.854460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T03:40:25.706499Z digest=sha256:e541b9705596d89552f285c81246b459121cfbf076c652e0662f91ecb878fa93

Observation a41acd6a-a1c2-4b54-8f11-fec289847cce · inbound

LLM-First Search: Self-Guided Exploration of the Solution Space cites this paper.

LLM-First Search: Self-Guided Exploration of the Solution Space On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:41.579717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:41.579717Z digest=sha256:636ef5a776939ba1c59ae2f118b392e616dd361442f1af3c395a39d04885406c

Observation 0ed703d5-4526-4335-9935-6d8a042fb6da · inbound

Reasoning-Finetuning Repurposes Latent Representations in Base Models cites this paper.

Reasoning-Finetuning Repurposes Latent Representations in Base Models On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:50:00.873671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:50:00.873671Z digest=sha256:a7adc7cd08843f34a72011150e1a309c524c20946ae5d939d80bce5c9f1e9f76

Observation 83d99c6c-72e6-4eff-8a06-18fbb0536bef · inbound

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete cites this paper.

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.178705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T15:48:48.046003Z digest=sha256:2f1f3bd22e2ae37d06c0d11d75cc9b7fa1e12fe8defaa87c2c6023125dce2ceb

Observation 7b5fd31f-e1af-469f-91b7-bb688a19f061 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 290

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.593548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:14bab6d3174d14dc5f25ee6aeb2e61e5b1114c3ea540e24053be98f49650b645