Pith. sign in

Paper Citation Record · LEDGER

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

As of 9 August 2026, this Paper Citation Record lists 100 of 217 outbound references and 1 inbound Pith citation observation for arXiv:2507.00004.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00004 v2

Coverage vector

measured 100 of 217 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:39.908059Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T12:29:24.439779Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:36:56.135904Z

Reference resolution

100 of 217 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ded0eeef-24b1-45e1-be13-4bc79ceb2429 · outbound

This paper cites Scaling Laws for Neural Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Neural Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.492720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.492720Z digest=sha256:a51255b5bbb6036afc9beb1d581c4a1acd7453decb2e6d0975d1e4f95bb52541

Observation bb086a52-8788-4114-a846-700cf8dbfb0e · outbound

This paper cites Training Compute-Optimal Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Compute-Optimal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.498613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.498613Z digest=sha256:de5dbc3bc2015b6802d3999558d1980dd2bce9c0ae8671bb319101ee2991beb1

Observation b0430e10-3c59-4534-997b-f8613924f1d5 · outbound

This paper cites Compute-Optimal LLMs Provably Generalize Better With Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Compute-Optimal LLMs Provably Generalize Better With Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.503440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.503440Z digest=sha256:6d3fdc82ea9ef2dcdd8918fa15989f47c1f134d169b39b9a782b5dedc225fefb

Observation f0d827f6-3399-4524-9224-f3b73361d949 · outbound

This paper cites Training compute of frontier AI models grows by 4-5x per year,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training compute of frontier AI models grows by 4-5x per year,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.507915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.507915Z digest=sha256:035fa519c04bc0b7c171dcc6ba7f2c7da5b1e2d4303c63179857b3c6135c14f6

Observation 5efef214-44c5-4040-8b6f-e7af7d39138a · outbound

This paper cites Increased Compute Efficiency and the Diffusion of AI Capabilities.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Increased Compute Efficiency and the Diffusion of AI Capabilities

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.512826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.512826Z digest=sha256:c6a38cbae3b2f847345bbd5ba4416f02a245b44755a762056b77b9935c9462b7

Observation 983dcfe7-b288-45fb-9d9d-5b300cfdfd01 · outbound

This paper cites Measuring the Algorithmic Efficiency of Neural Networks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring the Algorithmic Efficiency of Neural Networks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.517278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.517278Z digest=sha256:d46e0b8fcce1393648f56a91b90c3a03ccf0f3496d87366c9dc35df9f55764c2

Observation a40e8546-a6c2-4358-9f2c-4b0cc46c1a01 · outbound

This paper cites Algorithmic progress in language models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Algorithmic progress in language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.522279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.522279Z digest=sha256:b2ea57984007f8dc26fb95e3d0e5f521172e9b4982937d01109ca9c9d5be59a6

Observation 55527bc7-82f6-45e1-a4e7-de716a995acb · outbound

This paper cites Claude’s extended thinking,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Claude’s extended thinking,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.526454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.526454Z digest=sha256:af10eddc9bf2f3c10c902e86942e7da433e031f8fcd75ae015b0b0f2abbd08f1

Observation 94aeaad4-b318-4e1c-8996-82fa3d566d7b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.530848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.530848Z digest=sha256:8c69d3359021f8f9894d07a17e5a4e1cd0fea2b49c89c74e7de60173afdab14a

Observation d176cbbe-b1c1-4eeb-9c9f-b04f7c48f3cd · outbound

This paper cites Gemini 2.5: Our most intelligent AI model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Gemini 2.5: Our most intelligent AI model,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.534878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.534878Z digest=sha256:0e855e840a19638f3a07d83f301c20ca9795b8dd0e65d8008f564f3d8bf7f1a3

Observation d7f1571f-5cf1-4269-b93e-c61995d18531 · outbound

This paper cites IBM Granite 3.2: Reasoning, vision, forecasting and more,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search IBM Granite 3.2: Reasoning, vision, forecasting and more,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.538768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.538768Z digest=sha256:f7c0e80e5de75cd1d42385a838762beb5e77bc91826007dc731d6a45e53d4735

Observation be84ecfc-bee3-440e-87d9-358799f52e61 · outbound

This paper cites Phi-4-reasoning Technical Report.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Phi-4-reasoning Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.542544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.542544Z digest=sha256:76b86f0386da7acd63b8e56bcc644fdeb1f1f5a678dac2ceb0efac7490e2ec57

Observation 1fc145bf-976b-4c12-b63e-7551b5072a6c · outbound

This paper cites OpenAI o1 System Card,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search OpenAI o1 System Card,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.546640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.546640Z digest=sha256:612a69a38dde18fa53179dc09dbcc2d09173fca6af22dd0e05965cba3e08e267

Observation 0201b25d-8625-4501-8250-aa7f9357c89b · outbound

This paper cites Introducing OpenAI o3 and o4-mini,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Introducing OpenAI o3 and o4-mini,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.550360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.550360Z digest=sha256:54da149c219d4912558236cba5070da3eb444290f4428ebc96dc651ac3c9986b

Observation b1190000-cdb1-46d9-90f8-e9f8ffd15ccb · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Grok 3 Beta — The Age of Reasoning Agents,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.554243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.554243Z digest=sha256:208d812a7ffe508e52a7046d3745b48e6b23e8004a5c211860734d294daae510

Observation be22cfff-d3eb-4ee5-b11f-ad8d106bbf3b · outbound

This paper cites The growing energy footprint of artificial intelligence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The growing energy footprint of artificial intelligence,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.557955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.557955Z digest=sha256:3679ace2fa08b177163da4d73f22e5037794b07609ebfd2afbb69a5b0f1f8e1c

Observation 77f11932-1a2d-439a-9e21-ecd23f414404 · outbound

This paper cites Estimating the carbon footprint of BLOOM, a 176B parameter language model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Estimating the carbon footprint of BLOOM, a 176B parameter language model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.562176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.562176Z digest=sha256:ffe1b36729a338bc7963453955d8f563ac1daa172aab13001b3bc4bfbe3a10f8

Observation a8743a1b-2298-4a01-9984-c89d0c621898 · outbound

This paper cites The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.565848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.565848Z digest=sha256:062bec86a90191d11a6e8be7ebd0accd46886dd957510fd382519bd6a2fc04d9

Observation d5477b62-9f63-4959-9c94-04557d63739a · outbound

This paper cites Sustainable AI: Environmental implications, challenges and opportunities,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sustainable AI: Environmental implications, challenges and opportunities,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.570414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.570414Z digest=sha256:13ddf1d67ed0491637216f7db9cdb8cb20ce113969d57231672a819fc7745869

Observation 43c6e430-13ee-4da6-b7ea-0fefbd694ff9 · outbound

This paper cites The next wave of AI: Demand and adoption,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The next wave of AI: Demand and adoption,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.574308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.574308Z digest=sha256:ad0c5ea2734fed74db012c5dc754410b1227e0ae8a775939dfc0e495003dd8ab

Observation 96a5ee41-d563-4a26-8ec6-1a7235947640 · outbound

This paper cites From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.578208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.578208Z digest=sha256:f6f0b960231fb14b4384e8e87c4a6053fda9415a2ff9baefdc1e09d3e3659209

Observation bd8f6d77-36a2-4910-a6a6-dcf1bfd64aae · outbound

This paper cites an unresolved cited work.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.582584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.582584Z digest=sha256:d3e23f9ccb9071517be204f8d15f4c8ba2ae8d22450241f8e296496f5d933d00

Observation 4cb6bb14-8ced-428d-8814-8dfb901fea59 · outbound

This paper cites 1B user messages sent on ChatGPT every day,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search 1B user messages sent on ChatGPT every day,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.586504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.586504Z digest=sha256:76384aafebbb51cf871a9f9d7280214e5f5166dea7e78033e6bba6db7cd095e2

Observation b3a8d728-3b7d-46d4-879a-679fc2647aee · outbound

This paper cites ChatGPT added one million users in the last hour,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT added one million users in the last hour,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.590585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.590585Z digest=sha256:757a412325d09cfd06c7d5dee440583cb3045d05966bf2cbb55383ee0ace18a2

Observation 97a61648-86eb-4065-9d7d-d6e0b30c8fbf · outbound

This paper cites ChatGPT statistics and user trends (2025),.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT statistics and user trends (2025),

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.594451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.594451Z digest=sha256:e6856b10c3972a8d6c8998dc66d4613c4779a545b45f458eee04b63d84a78a47

Observation c6daaac8-50e7-423f-b605-e0e4e0a7907d · outbound

This paper cites A systematic review of Green AI,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A systematic review of Green AI,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.598332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.598332Z digest=sha256:5931295fbffafefc8755af8a4df8ec95d027051a6b71a2a87ef294fe89abcd25

Observation 80a688e5-b70a-40ea-91c4-956cb869db40 · outbound

This paper cites Deep Blue,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Blue,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.602174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.602174Z digest=sha256:fa49caf613a4a13ee012e30177a20ac0aba515f6a922b701333e3630e9acfaa5

Observation 6e7aebae-75b6-4a46-b0a6-9400c32231a5 · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go with deep neural networks and tree search,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.606172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.606172Z digest=sha256:bc06954dbe6110ef244e1a2d3160d324c1a1c2dff63975179a2f2bcf832a82d3

Observation e5c0c3f8-e05a-491c-b108-54348491f428 · outbound

This paper cites Mastering the game of Go without human knowledge,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go without human knowledge,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.611183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.611183Z digest=sha256:39f5cc7c29aee47249a271d3e45f7d490d878ace8a72b9a6ce5194944408f27c

Observation 3d9dc32f-db41-44a7-8c75-c83962b8fee6 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.615289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.615289Z digest=sha256:8dfdf15562f37464d6ae940ab3f3173197b72a04028e73bb9ddfc14f6a1485b3

Observation 6f8b12d7-0413-43ee-87ec-b1f2fa7277c0 · outbound

This paper cites Scaling Scaling Laws with Board Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Scaling Laws with Board Games

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.619445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.619445Z digest=sha256:1c93ebc9d0b7247db601c83dfc6307df6894db42d6f17173a16cff341f66db00

Observation 638d68f3-0448-4836-bc9a-b4bb22aa41b0 · outbound

This paper cites Safe and Nested Subgame Solving for Imperfect-Information Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Safe and Nested Subgame Solving for Imperfect-Information Games

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.624213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.624213Z digest=sha256:8c1a444c9915f752be47c2aad2b9ee79dba27cb1c76551fcb3e590c1cca4404d

Observation a549f08b-343c-4022-aca5-e6358296ec98 · outbound

This paper cites Human-level play in the game of Diplomacy by combining language models with strategic reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Human-level play in the game of Diplomacy by combining language models with strategic reasoning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.628464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.628464Z digest=sha256:d6bbba3334be4df22c60799a3985c105882d347338ed2e4ebcb09f9092f84178

Observation c111cbba-fc11-4128-a660-941e7adc9728 · outbound

This paper cites Emergent Abilities of Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Emergent Abilities of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.632468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.632468Z digest=sha256:5df652cd808c345ee442f19b234edd24f298f5dd79974491491fe64006061e18

Observation f63715f7-46e5-4879-84ca-a789399dc89d · outbound

This paper cites An information theory of compute-optimal size scaling, emergence, and plateaus in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search An information theory of compute-optimal size scaling, emergence, and plateaus in language models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.636674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.636674Z digest=sha256:77fb3db57fd725d9bf6eeaee883378a24a822dd60703a4681dbd6d83d4ece999

Observation d0251886-4a32-4390-8104-c926f87aa3af · outbound

This paper cites Multi-task Language Understanding on MMLU Leaderboard,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Multi-task Language Understanding on MMLU Leaderboard,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.640705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.640705Z digest=sha256:b5122a877bf1b245aa52012ca0df79e3f91ca40f3370758d24733bdf7dbc5d63

Observation 29c7dd85-4868-465f-8763-e1222da9e0d5 · outbound

This paper cites Measuring massive multitask language understanding,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring massive multitask language understanding,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.644826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.644826Z digest=sha256:a4aa3fb5cae2d9c3b33299356b5cf250982d963f111b37d3019fa2c5dc988c93

Observation 2c45d034-17db-427c-a26d-ef2b36bfdb98 · outbound

This paper cites Are emergent abilities of large language models a mirage?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Are emergent abilities of large language models a mirage?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.648721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.648721Z digest=sha256:55cf22df8e2da2986de86f9fc4c3535b09a5295a8e0f5bba1b9395cbc6328c32

Observation e0d401d8-160a-41a9-9e0f-bb25ed697939 · outbound

This paper cites The quantization model of neural scaling,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The quantization model of neural scaling,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.652682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.652682Z digest=sha256:cafd513bb8d010944cc1172e767c271fd38a59997aa4f0ee34916ae209b5a86b

Observation e6616e3c-20cf-4f50-8c31-bc0dc045550a · outbound

This paper cites Circuit tracing: Revealing computational graphs in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Circuit tracing: Revealing computational graphs in language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.656514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.656514Z digest=sha256:16d363385d74d7eac80162e0817eef32ef14c86f69275d7403369e8c923d6079

Observation 3fa94251-d92d-4778-9dc5-ae14f5e8ac3f · outbound

This paper cites Curriculum learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Curriculum learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.664621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.664621Z digest=sha256:bd5869bf0a293119e8f4c7883524ac2ccdba7d74f9dedf0bca0135952260c3e6

Observation 86c4006f-a309-4535-b9cf-08632f0485c9 · outbound

This paper cites A Theory for Emergence of Complex Skills in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Theory for Emergence of Complex Skills in Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.668555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.668555Z digest=sha256:22bc96f3b7ad1e8e9ecccf060dd0a741a0b4eb63831485de080d29b7ab8087e5

Observation c12e0140-4c77-4313-8279-37d316de2be2 · outbound

This paper cites A mathematical theory for learning semantic languages by abstract learners,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A mathematical theory for learning semantic languages by abstract learners,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.673016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.673016Z digest=sha256:a951e754cb50dbff39de695761b66ffbfb032a75b1afdca831cc6a55efebd7de

Observation 66468c21-ea12-4cdf-83f6-61997079e119 · outbound

This paper cites Skill-Mix: a flexible and expandable family of evaluations for AI models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Skill-Mix: a flexible and expandable family of evaluations for AI models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.676978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.676978Z digest=sha256:f362fd5c66bbe51e2d8ebfefe055affa7209ae9129f390cec033b302fb2ca289

Observation 0cd33ccb-1cbf-4e55-ad77-ba6e7bff0b80 · outbound

This paper cites The learning curve: implications of a quantitative analysis,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The learning curve: implications of a quantitative analysis,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.680836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.680836Z digest=sha256:c3aa79b94e2574d3a97dca94618800517706b6379830d121a9781c608a50aabe

Observation 27c29121-cbce-4bce-83b4-97c2df0c216f · outbound

This paper cites Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.684918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.684918Z digest=sha256:6c6d5995b9fa39175d29d2e834e68d8b3e8321c4e1cf7c43eff63e9dedf7e0a6

Observation 91c8e278-8ca1-4dfa-927a-b870f9af7cd7 · outbound

This paper cites A first-principles mathematical model integrates the disparate timescales of human learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A first-principles mathematical model integrates the disparate timescales of human learning,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.688862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.688862Z digest=sha256:e1a7ac2c8a21b05dc2e194e77ad125419061a902dff719c4cb9236e92313250c

Observation 6bd50324-3c6f-4f18-971b-9b663a5a8c9f · outbound

This paper cites Spin-glass models as error-correcting codes,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Spin-glass models as error-correcting codes,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.693324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.693324Z digest=sha256:5d8e5b297fb88c4f8bee3ed9c23f57ceb4862918434f08e8efefb47efd6ee037

Observation 8491fe46-a0cd-40d9-9638-2b9ad7f199dd · outbound

This paper cites Newell,Unified Theories of Cognition.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Newell,Unified Theories of Cognition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.697193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.697193Z digest=sha256:d7e4c25aa63767c2c5bf52c8df186e6ff944424a9fa795822bf5d8543aa0f775

Observation 4c225486-1177-463a-bbeb-f745387c52fb · outbound

This paper cites Barab ´asi,Network Science.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Barab ´asi,Network Science

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.701454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.701454Z digest=sha256:8871296d712fac62758f3013701e62dc02511b771e076468ba06506cde204fb6

Observation b721141f-d0e6-4ec5-b9df-f670f1449ae9 · outbound

This paper cites Learning curves: Asymptotic values and rate of convergence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Learning curves: Asymptotic values and rate of convergence,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.705522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.705522Z digest=sha256:9f2ae5ff2637c5814a63d8c680f9b8103af1d3e2459a45574bfe689315031f00

Observation 27dac46d-595c-4af0-be03-16caa56f7f4b · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Learning Scaling is Predictable, Empirically

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.709672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.709672Z digest=sha256:45fded595f6f94b533f0e3d9a7051a696acb6e2fbf1e583d628ec551db62f819

Observation 77e77b35-f651-4a00-96d2-619bda196821 · outbound

This paper cites A Constructive Prediction of the Generalization Error Across Scales.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Constructive Prediction of the Generalization Error Across Scales

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.713752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.713752Z digest=sha256:4a76636c21919d303bcb4018eb924a2468e1e0d39a3735cb0f08acb57c1cbb31

Observation b2483da6-1c56-4ecf-b3e0-8f16bcb95355 · outbound

This paper cites Language Models are Few-Shot Learners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models are Few-Shot Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.718254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.718254Z digest=sha256:735f8835d273187f28496b19ab42fd0cbf49ef94fcbbdaa3d64b49f20a7ba337

Observation a5a4d548-51d1-4c27-b5a3-208835a5ce7a · outbound

This paper cites Prediction and entropy of printed English,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Prediction and entropy of printed English,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.722424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.722424Z digest=sha256:aafeef6060d4eac1290606c26fcf54daa3766378e3ff11cf878fa524a02d1396

Observation 6366e3eb-1eff-489f-9840-95594bbd5bf6 · outbound

This paper cites Explaining neural scaling laws,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Explaining neural scaling laws,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.726859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.726859Z digest=sha256:c706a294110ca5dd3e5f5970f44414df914a4506a3984d20ea8046b5e97eefdc

Observation b54a65a2-cecf-44f0-8acc-597adfd2fe59 · outbound

This paper cites Towards a universal scaling law of LLM training and inference,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Towards a universal scaling law of LLM training and inference,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.730817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.730817Z digest=sha256:cece514e65bdf3997e394168ee31b70c2e22fffcebedb416c66ee0511d08c1a2

Observation 1e02fe25-2a3f-4352-8ea8-8dba24901087 · outbound

This paper cites The Llama 3 Herd of Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Llama 3 Herd of Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.734911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.734911Z digest=sha256:0a3bb402e7d81e1478a40dd3230fe2ffd2bc9e335740febc66a46aa510cb5736

Observation 7786d76e-84fb-4402-a578-d68742c0b14c · outbound

This paper cites Densing Law of LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Densing Law of LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.739462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.739462Z digest=sha256:82d9f3e495a248155d997af2c4573f74e04b137c5f5c196f145ba05c86e287a3

Observation 341e967e-f53c-44d0-b95a-640968e74a98 · outbound

This paper cites How predictable is language model benchmark performance?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search How predictable is language model benchmark performance?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.743953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.743953Z digest=sha256:b1ed9d5368a5a75f08a3c7b3ecec673d8e7e8396caf2c8a8de0086461de607f6

Observation 55d90686-d264-4dcc-acd3-a88ee28203c9 · outbound

This paper cites Observational Scaling Laws and the Predictability of Language Model Performance.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Observational Scaling Laws and the Predictability of Language Model Performance

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.748185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.748185Z digest=sha256:969228faa912f681df1dfeb28968a40ad8834e7e3a7b03f473b32b03e922cacb

Observation 7212905d-590a-4677-bc73-d486d3d50ec8 · outbound

This paper cites Language models scale reliably with over-training and on downstream tasks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language models scale reliably with over-training and on downstream tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.752393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.752393Z digest=sha256:9e29aa57bc308ec59e805190749a3ac323556954b248b54bf1926bca7b2c1ddb

Observation f1e3447c-cb80-4840-93aa-5ed3ab6d4741 · outbound

This paper cites Broken Neural Scaling Laws.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Broken Neural Scaling Laws

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.756787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.756787Z digest=sha256:78f6c8bbf46aa20023051e79087460bbfbfba84b0273c8a496a9b3d6f4d1a72b

Observation 76fd424c-0c4b-41c9-8969-76a406233349 · outbound

This paper cites Scaling laws for downstream task performance in machine translation,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling laws for downstream task performance in machine translation,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.761555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.761555Z digest=sha256:b84b9631aac063fa95935a67e0119ee5404b2b2f7dc240283107fb87340b6796

Observation 24a0ee2a-be06-4e0c-ab17-dac540d80973 · outbound

This paper cites Exploring the Limits of Large Scale Pre-training.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Exploring the Limits of Large Scale Pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.765677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.765677Z digest=sha256:d8681f8da9f9a2475333a6267d50ab7b699b6a98659cadd430886bb5cd777420

Observation 0065568d-23ea-46ab-955f-c9b2bee4b5c9 · outbound

This paper cites Scaling Laws Do Not Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws Do Not Scale

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.769818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.769818Z digest=sha256:fae2e74778c53446f0c4de61eb849d0f99d4fc4114b081f1a65591a232b99da7

Observation 663ecb2d-e843-4e8e-8586-d528424257a7 · outbound

This paper cites Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.773993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.773993Z digest=sha256:ff69d30bc2f8945048905680b5df5475abe66723cb67a2040a6d4211710291f5

Observation 798b2385-8575-4476-9590-01cbf7243f93 · outbound

This paper cites Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.778229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.778229Z digest=sha256:f53714356325dc2a171c2ca4b9298afb66f5f83c8873b151b1be7ee78e1a332e

Observation 29406675-cb2b-4863-b3dd-2c3dd95f3e92 · outbound

This paper cites Overtrained Language Models Are Harder to Fine-Tune.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Overtrained Language Models Are Harder to Fine-Tune

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.782404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.782404Z digest=sha256:6b5d9dda42015cb45f87d2320feee5bf0b1e7a2e08d61c539ab8429b5af2e86d

Observation 4638aa3f-b743-452f-b280-1e9c4233a83c · outbound

This paper cites Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.786753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.786753Z digest=sha256:132601f2435cdfa3d800d8de8bb8c7bdca24cc78d506e41ef59a9af70a36f099

Observation e9dd4514-ab59-4454-848f-e22edde220ed · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.791105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.791105Z digest=sha256:ce15c1363050f1015cde62eabd3639b3c00ad01d3e7506ec2263de0d6515a6b3

Observation 33d1bc75-12dc-4f7f-984d-dfbfd9f856a9 · outbound

This paper cites BBQ: A Hand-Built Bias Benchmark for Question Answering.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search BBQ: A Hand-Built Bias Benchmark for Question Answering

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.795184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.795184Z digest=sha256:f1cf2e4227ed2558466403a4a03fb339bc13eb4dd9017afa110d553f688f69e9

Observation 50259960-870e-45ad-b591-0b717961a115 · outbound

This paper cites Inverse scaling can become U-shaped.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Inverse scaling can become U-shaped

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.799434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.799434Z digest=sha256:b11588c6cf93266fb6f9c2421fe24fd2c7a14fb8ae5d89bbd65a2c5715220449

Observation b52716c6-d685-48ae-a9a6-3271f62f6f1d · outbound

This paper cites s1: Simple test-time scaling.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search s1: Simple test-time scaling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.803725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.803725Z digest=sha256:cdd7b756eb2af0ae3bbd747bd677de4511e7fa452ab5504e14583492fe80a332

Observation 5ad86b0f-93bd-4328-8700-4bce37b0ea1c · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.807957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.807957Z digest=sha256:a39e91334fb71867fec1a892835f4697f5edb33145ba28b56b42b837389c866d

Observation 5cbeb825-46fa-49b6-809b-3a8ec04067d0 · outbound

This paper cites Large Language Models are Zero-Shot Reasoners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Large Language Models are Zero-Shot Reasoners

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.812191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.812191Z digest=sha256:0e3a05c720716251e41f5a17da34e304c302d9e7c1362a4ddf8c8f31728f9dc8

Observation 2160b047-c53e-483c-8455-5b6271a68dcb · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.815981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.815981Z digest=sha256:552eb9a6d10b175b69436d6510dfc9be684b0f2861d66a6a15d5df9bbf4601e0

Observation 6eb5c544-c2dc-4356-ae67-c3edea4480a2 · outbound

This paper cites Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.820330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.820330Z digest=sha256:25e1e05c5c0a1ca3e4c682be377c8fa8d22616c7137d8cb6367fdc53bfd2866f

Observation 395dad7c-5bf2-4692-9418-3263c1e2e7c9 · outbound

This paper cites Sequence to sequence learning with neural networks: What a decade,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sequence to sequence learning with neural networks: What a decade,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.824523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.824523Z digest=sha256:0756c0b712242bd1846125a8f61752a0241c3c8146ffb33b906b511c5ae39c1e

Observation 43d03769-8801-4aad-8193-17701a25af25 · outbound

This paper cites AI doom from an LLM-plateau-ist perspective,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI doom from an LLM-plateau-ist perspective,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.828440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.828440Z digest=sha256:d44e379bb014674f76a5b543777414d50ca9564b1e3c01e439c635957d59fac3

Observation dd5603ca-4f42-4b38-a44e-be59e9426aa4 · outbound

This paper cites The first wave of AI innovation is over. here’s what comes next,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The first wave of AI innovation is over. here’s what comes next,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.832462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.832462Z digest=sha256:c1a0828feab4be80f6b1e99fd576f9b9781bd5bb27432b1d5ceadc27921f0ca5

Observation 0a53bca0-df12-4823-90de-32112cda7561 · outbound

This paper cites AI won’t plateau — if we give it time to think,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI won’t plateau — if we give it time to think,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.836408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.836408Z digest=sha256:29a6c64d4c6bc261f2ca7a283a5082900912f46d1343a52d7e52937a1ef582c2

Observation 22f4293a-5b65-424e-ba50-474d1949f1d5 · outbound

This paper cites Scaling data-constrained language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling data-constrained language models,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.840145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.840145Z digest=sha256:f65e126d6b6043e46df02fd80be763e23daefc6be88fe75d5773caa22e70cfc1

Observation 023ee840-7c37-4322-992c-200a388fecf4 · outbound

This paper cites Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.844389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.844389Z digest=sha256:94810f1c83f23cffedb06d2570f0a1938526c15a18e24fb5c52b190d92a36633

Observation c6de41cc-bec8-40c6-bcd7-1aface91964a · outbound

This paper cites The rising costs of training frontier AI models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The rising costs of training frontier AI models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.848494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.848494Z digest=sha256:7e155df00c015c5046302931bff98a741615a6da6403955227487ff24b11a8a3

Observation 2fbe18fe-d46e-40fe-be04-734db104bfbf · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.852434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.852434Z digest=sha256:23536a4d063266312330f22bfe7539b690b8e5d2ec441cb12672731498953f81

Observation 8a60c330-fc02-4424-aa94-100361efb22b · outbound

This paper cites Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.856507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.856507Z digest=sha256:8dc602f08c2dd5617d5945621ccbb20b77af62840b46f284a3108902c196de52

Observation caa6192b-6ee1-41e8-a587-77f1a904d156 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Efficient Streaming Language Models with Attention Sinks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.860710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.860710Z digest=sha256:95f3b480abd59b2071fe03c055d305833418026d32e885c4902475e7b57c3b94

Observation 71ffee8b-c47a-4cfb-a28b-3ff54ef3c5ad · outbound

This paper cites Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.864758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.864758Z digest=sha256:698c57d728413fe6ae86b5784bb4820f19d9fa8bb861994dab4c92c1f3b739bb

Observation 489eb03e-4f65-440f-b1aa-4fd2ce6adefb · outbound

This paper cites The curse of CoT: On the limitations of chain-of-thought in in-context learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The curse of CoT: On the limitations of chain-of-thought in in-context learning,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.868843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.868843Z digest=sha256:e721484c421c068a22b11e9f76c0b41dd54b0b4970b24c6ef7fd399b2861b888

Observation 634eb7db-c73e-4d8a-92e9-9d8456bc662d · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.872688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.872688Z digest=sha256:a59e496f8d72d740ee6fd794426d6b35a8b9b5bfc441eb7680ec4897a32752c6

Observation 79cc6bdc-7dea-481c-bbe1-42664f1d2e42 · outbound

This paper cites Let's Verify Step by Step.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Let's Verify Step by Step

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.876992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.876992Z digest=sha256:e1e4e23c8963e45a09effec10648db319bac55950aab9c82709b5c80e655ef42

Observation a884701b-03ed-4fe4-9a68-c650ccb5ee85 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Verifiers to Solve Math Word Problems

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.880977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.880977Z digest=sha256:ce4abb62921d3bbf364fad29aaa0d19bdab38009a9bdfd76721d068e3424a5cd

Observation e262f215-9b8c-4ac1-88db-db32dce362a4 · outbound

This paper cites When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.884774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.884774Z digest=sha256:b91ee5bfe3fa00af0d6b811bde69e3994822bc919d6b7ae7d28377ce1024d00e

Observation f052279f-279d-41c7-83a1-2efabc1fa6e6 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.888499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.888499Z digest=sha256:670e0edfa6885a89835be8a05cdad5bb60423e47b34af1504108ecdbdc5a30ce

Observation bf2247ff-05b5-4065-9db4-a24a122bd5a1 · outbound

This paper cites Solving quantitative reasoning problems with language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Solving quantitative reasoning problems with language models,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.892256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.892256Z digest=sha256:5bc6c876eae0cc1f552a198ccc39a652526d793341c889003b15babd38f7901d

Observation 055e7f58-e5b6-429d-bf48-6546b15f8b3d · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Refine: Iterative Refinement with Self-Feedback

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.895889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.895889Z digest=sha256:c1aa139e8b23aecd8fdfbe43bff13a981dbd5d64a89022c51661a4baf3f1c84b

Observation bf4a5fc3-212a-4143-b14f-e9186f1fb833 · outbound

This paper cites Cost-of-Pass: An economic framework for evaluating language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Cost-of-Pass: An economic framework for evaluating language models,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.899967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.899967Z digest=sha256:e58a83cf7f8b9610adf35c3f64f934e8ab1f3c56c1e32e8a99136a472bf330bc

Observation 2aa08e6c-7e5a-4790-be4b-c704137801f3 · outbound

This paper cites Scaling Laws for Transfer.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Transfer

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.903603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.903603Z digest=sha256:f99931314ec9ad88d8981062a8d7f6669590b544501211d98881681c8ee3be51

Observation 45d542b2-0c07-4f17-ba0f-eb4f23d3ebc4 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reproducible scaling laws for contrastive language-image learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.908059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.908059Z digest=sha256:155ee3d9485f15f76ba8d759667f40d41aa4b942592fb09a2ae39c6661eeb385

Pith citing papers

Observation 5e5e584d-abf1-49fc-9722-19276522f94e · inbound

Two AI Metrics Diverged: Will it Make All the Difference? cites this paper.

Two AI Metrics Diverged: Will it Make All the Difference? A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:56.137222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-02T12:29:24.439779Z digest=sha256:1f5dfdb06d244155c8fc2be76bdeb7f1b4cad24f7d0aaf048642eb1bc49a4057