Pith. sign in

Paper Citation Record · LEDGER

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

As of 16 August 2026, this Paper Citation Record lists 100 of 217 outbound references and 1 inbound Pith citation observation for arXiv:2507.00004.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00004 v2

Coverage vector

measured 100 of 217 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:39.908059Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-02T12:29:24.439779Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:36:56.135904Z

Reference resolution

100 of 217 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ded0eeef-24b1-45e1-be13-4bc79ceb2429 · outbound

This paper cites Scaling Laws for Neural Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Neural Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.492720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.492720Z digest=sha256:f338aaddaafbc536c9deb417cf879a6a97f350c571d9267a9df289d7cfbe8e3d

Observation bb086a52-8788-4114-a846-700cf8dbfb0e · outbound

This paper cites Training Compute-Optimal Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Compute-Optimal Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.498613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.498613Z digest=sha256:5be286f04d5118e3c4de1d9f74e4ac55e0f6a34f8ba1336f2dce2056b62b5add

Observation b0430e10-3c59-4534-997b-f8613924f1d5 · outbound

This paper cites Compute-Optimal LLMs Provably Generalize Better With Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Compute-Optimal LLMs Provably Generalize Better With Scale

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.503440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.503440Z digest=sha256:57e08501187549ee8532e8909976720ad00fc2c5ddc149f6afa788e8da12ed1f

Observation f0d827f6-3399-4524-9224-f3b73361d949 · outbound

This paper cites Training compute of frontier AI models grows by 4-5x per year,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training compute of frontier AI models grows by 4-5x per year,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.507915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.507915Z digest=sha256:732d3c2d5576171b93142b38742b9555f2559a661fda4bd0276559f0d81125e9

Observation 5efef214-44c5-4040-8b6f-e7af7d39138a · outbound

This paper cites Increased Compute Efficiency and the Diffusion of AI Capabilities.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Increased Compute Efficiency and the Diffusion of AI Capabilities

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.512826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.512826Z digest=sha256:794b5498bdcb28bf9d215999b236841acdb656601a94d56f0494c3fa16b30712

Observation 983dcfe7-b288-45fb-9d9d-5b300cfdfd01 · outbound

This paper cites Measuring the Algorithmic Efficiency of Neural Networks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring the Algorithmic Efficiency of Neural Networks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.517278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.517278Z digest=sha256:61e13e15bc7340bb06bffac1e06ec06bb519bc0c43a35c971489f06a1d4482b3

Observation a40e8546-a6c2-4358-9f2c-4b0cc46c1a01 · outbound

This paper cites Algorithmic progress in language models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Algorithmic progress in language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.522279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.522279Z digest=sha256:87dd5dce71c392600618e5430689e09bdca227308a41ddd50579d841a6d3851b

Observation 55527bc7-82f6-45e1-a4e7-de716a995acb · outbound

This paper cites Claude’s extended thinking,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Claude’s extended thinking,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.526454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.526454Z digest=sha256:de3f7febafc7da7cca272dcf5c26e674c8f65fe5670704def08fc03a39de2597

Observation 94aeaad4-b318-4e1c-8996-82fa3d566d7b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.530848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.530848Z digest=sha256:3c48de300e14d1513ba6c7ed84f4c76315250d386de355dfe7ff43976adf0834

Observation d176cbbe-b1c1-4eeb-9c9f-b04f7c48f3cd · outbound

This paper cites Gemini 2.5: Our most intelligent AI model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Gemini 2.5: Our most intelligent AI model,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.534878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.534878Z digest=sha256:590ff2882b20b307639927484a82398222a2bcb32de8c8671b607dba19fe818d

Observation d7f1571f-5cf1-4269-b93e-c61995d18531 · outbound

This paper cites IBM Granite 3.2: Reasoning, vision, forecasting and more,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search IBM Granite 3.2: Reasoning, vision, forecasting and more,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.538768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.538768Z digest=sha256:97c8db6df815a83222561a9f4911420478f162e34b12d5ed76c39087fd280306

Observation be84ecfc-bee3-440e-87d9-358799f52e61 · outbound

This paper cites Phi-4-reasoning Technical Report.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Phi-4-reasoning Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.542544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.542544Z digest=sha256:4de2d184f5287855b4cb3080f295d8d020e84d0531c99fdd08e4042d7eae6e03

Observation 1fc145bf-976b-4c12-b63e-7551b5072a6c · outbound

This paper cites OpenAI o1 System Card,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search OpenAI o1 System Card,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.546640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.546640Z digest=sha256:5f7ce23a2a2e094ac265c2723deab59038bddf6a9a6d8157ba1ec781439ede68

Observation 0201b25d-8625-4501-8250-aa7f9357c89b · outbound

This paper cites Introducing OpenAI o3 and o4-mini,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Introducing OpenAI o3 and o4-mini,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.550360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.550360Z digest=sha256:6ec86c1b2861859f9cee1435227d3dcf939036715639c8627a59b3341611ec2f

Observation b1190000-cdb1-46d9-90f8-e9f8ffd15ccb · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Grok 3 Beta — The Age of Reasoning Agents,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.554243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.554243Z digest=sha256:93d27a33e74a1bc618b2caed3fa7ed24493b0630c4fdba5f845249430c36a37c

Observation be22cfff-d3eb-4ee5-b11f-ad8d106bbf3b · outbound

This paper cites The growing energy footprint of artificial intelligence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The growing energy footprint of artificial intelligence,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.557955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.557955Z digest=sha256:05430769d543b094ebd5e4c7ac348f52a1331133ca27d9296d55b694eeec7a34

Observation 77f11932-1a2d-439a-9e21-ecd23f414404 · outbound

This paper cites Estimating the carbon footprint of BLOOM, a 176B parameter language model,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Estimating the carbon footprint of BLOOM, a 176B parameter language model,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.562176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.562176Z digest=sha256:8b66eeade456e086291e6761802c7e8b1dd72646a7eb765433bf8580165da026

Observation a8743a1b-2298-4a01-9984-c89d0c621898 · outbound

This paper cites The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Carbon Footprint of Machine Learning Training Will Plateau, Then Shrink

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.565848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.565848Z digest=sha256:2795c884a4200476a0321a1bb70a773d0e32e2df2c7166585403d25a48dca05d

Observation d5477b62-9f63-4959-9c94-04557d63739a · outbound

This paper cites Sustainable AI: Environmental implications, challenges and opportunities,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sustainable AI: Environmental implications, challenges and opportunities,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.570414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.570414Z digest=sha256:79f3376fa273bd5dc6f974c44fd5b20a4a307df479ea66f4510aad10940e5e8e

Observation 43c6e430-13ee-4da6-b7ea-0fefbd694ff9 · outbound

This paper cites The next wave of AI: Demand and adoption,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The next wave of AI: Demand and adoption,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.574308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.574308Z digest=sha256:87a84917e19474a1920cf07f4c3bfdb6a2cea640624cdf1283f1c6134f49adce

Observation 96a5ee41-d563-4a26-8ec6-1a7235947640 · outbound

This paper cites From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search From Efficiency Gains to Rebound Effects: The Problem of Jevons' Paradox in AI's Polarized Environmental Debate

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.578208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.578208Z digest=sha256:d659f374e1047f9ee52a07ebd608bfc422e36894725c82378be6ccc881adef7d

Observation bd8f6d77-36a2-4910-a6a6-dcf1bfd64aae · outbound

This paper cites an unresolved cited work.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.582584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.582584Z digest=sha256:f02a8321dedeeb3bef8712cb4a811b7999ef5372f98bf10498119be575f51115

Observation 4cb6bb14-8ced-428d-8814-8dfb901fea59 · outbound

This paper cites 1B user messages sent on ChatGPT every day,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search 1B user messages sent on ChatGPT every day,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.586504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.586504Z digest=sha256:3156029c2c4122fe4674a98ea647e4811eaba49b6a64cbaacef598f8edfa1bd7

Observation b3a8d728-3b7d-46d4-879a-679fc2647aee · outbound

This paper cites ChatGPT added one million users in the last hour,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT added one million users in the last hour,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.590585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.590585Z digest=sha256:382296e442238a723e567578775f89332981d9a92a3b4b457e8ef6fca74bfb0f

Observation 97a61648-86eb-4065-9d7d-d6e0b30c8fbf · outbound

This paper cites ChatGPT statistics and user trends (2025),.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search ChatGPT statistics and user trends (2025),

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.594451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.594451Z digest=sha256:cf41d4103137773d812eb103288f65389da323a85f000498c11852e3fd1f6876

Observation c6daaac8-50e7-423f-b605-e0e4e0a7907d · outbound

This paper cites A systematic review of Green AI,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A systematic review of Green AI,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.598332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.598332Z digest=sha256:a09f76cae19e6e531cd043e82f3e4f19e58782c4e31e294909e1f8c8d5cba73e

Observation 80a688e5-b70a-40ea-91c4-956cb869db40 · outbound

This paper cites Deep Blue,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Blue,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.602174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.602174Z digest=sha256:46b211799e2da0c80d822c931575586627a6ddfc5091ab454f66f60d6c98cce2

Observation 6e7aebae-75b6-4a46-b0a6-9400c32231a5 · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go with deep neural networks and tree search,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.606172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.606172Z digest=sha256:962cc6698aadf0478ee0f5ea2950ea5cb3e5204f0e4618527a59df8a00908763

Observation e5c0c3f8-e05a-491c-b108-54348491f428 · outbound

This paper cites Mastering the game of Go without human knowledge,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering the game of Go without human knowledge,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.611183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.611183Z digest=sha256:8a1ccdf80ae714852b586b2b66fcc5c0c08691c92384e8e0a028e904e8efa172

Observation 3d9dc32f-db41-44a7-8c75-c83962b8fee6 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.615289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.615289Z digest=sha256:4b8ac9a7084aa8e2efd74a595beb192e907e92b55bfba167808e8c299e2454ac

Observation 6f8b12d7-0413-43ee-87ec-b1f2fa7277c0 · outbound

This paper cites Scaling Scaling Laws with Board Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Scaling Laws with Board Games

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.619445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.619445Z digest=sha256:4e1fc2a302e39916f8c11f3f7d0edd12c0578bbd22a9a48378f52ae6a14f8d95

Observation 638d68f3-0448-4836-bc9a-b4bb22aa41b0 · outbound

This paper cites Safe and Nested Subgame Solving for Imperfect-Information Games.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Safe and Nested Subgame Solving for Imperfect-Information Games

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.624213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.624213Z digest=sha256:f0ddf748cc1321450e740d46304cd394db317a23e4747968aa823d84e3c3bdec

Observation a549f08b-343c-4022-aca5-e6358296ec98 · outbound

This paper cites Human-level play in the game of Diplomacy by combining language models with strategic reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Human-level play in the game of Diplomacy by combining language models with strategic reasoning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.628464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.628464Z digest=sha256:bc98a856b8b37c6ee2e538844a770c8a6c3bbddf0edf503aeb1e001714d83092

Observation c111cbba-fc11-4128-a660-941e7adc9728 · outbound

This paper cites Emergent Abilities of Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Emergent Abilities of Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.632468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.632468Z digest=sha256:4a9220db245337697f6c377af9c0e23ce39bc1d59828baff1d4a6b51f7981add

Observation f63715f7-46e5-4879-84ca-a789399dc89d · outbound

This paper cites An information theory of compute-optimal size scaling, emergence, and plateaus in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search An information theory of compute-optimal size scaling, emergence, and plateaus in language models,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.636674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.636674Z digest=sha256:2a3fdec7f1a75a0654be1c024ae648e2be015f66f5984224715eb82f695bc1ee

Observation d0251886-4a32-4390-8104-c926f87aa3af · outbound

This paper cites Multi-task Language Understanding on MMLU Leaderboard,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Multi-task Language Understanding on MMLU Leaderboard,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.640705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.640705Z digest=sha256:79fa8320f21d62e8a5d8817b5603eb6a8156715410bb09b78386ffc7e2372519

Observation 29c7dd85-4868-465f-8763-e1222da9e0d5 · outbound

This paper cites Measuring massive multitask language understanding,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Measuring massive multitask language understanding,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.644826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.644826Z digest=sha256:9b612af139abf22f32cfa42f1804d3d48f6c4938c08876a79d71db275131f974

Observation 2c45d034-17db-427c-a26d-ef2b36bfdb98 · outbound

This paper cites Are emergent abilities of large language models a mirage?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Are emergent abilities of large language models a mirage?

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.648721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.648721Z digest=sha256:f4d0983e0379200970f35529e4964597dab36e4e170f4c42349c96a9b3ade93b

Observation e0d401d8-160a-41a9-9e0f-bb25ed697939 · outbound

This paper cites The quantization model of neural scaling,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The quantization model of neural scaling,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.652682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.652682Z digest=sha256:061d0af9997b7b0667680d6027c598f7d91af68e6bbcf353c2d311f150a1f4a4

Observation e6616e3c-20cf-4f50-8c31-bc0dc045550a · outbound

This paper cites Circuit tracing: Revealing computational graphs in language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Circuit tracing: Revealing computational graphs in language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.656514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.656514Z digest=sha256:b748014d23eb8a34e98562574ecaf9129c6168dab1cd2054a3dde73405caaa72

Observation 3fa94251-d92d-4778-9dc5-ae14f5e8ac3f · outbound

This paper cites Curriculum learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Curriculum learning,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.664621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.664621Z digest=sha256:91e6600eaad209659a407b97a43849ba04187027bbf697755be8c1196f353a78

Observation 86c4006f-a309-4535-b9cf-08632f0485c9 · outbound

This paper cites A Theory for Emergence of Complex Skills in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Theory for Emergence of Complex Skills in Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.668555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.668555Z digest=sha256:c7dcc805cd9511f92746cca53570b52ac02f458ff5fd8d2be709d4ac7b9ca98b

Observation c12e0140-4c77-4313-8279-37d316de2be2 · outbound

This paper cites A mathematical theory for learning semantic languages by abstract learners,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A mathematical theory for learning semantic languages by abstract learners,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.673016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.673016Z digest=sha256:6584989a53fa0e69c43f88b34b8c748d713fa04750d63a52a5d17988ac805ae2

Observation 66468c21-ea12-4cdf-83f6-61997079e119 · outbound

This paper cites Skill-Mix: a flexible and expandable family of evaluations for AI models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Skill-Mix: a flexible and expandable family of evaluations for AI models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.676978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.676978Z digest=sha256:391b61bd9b589cb1f7436058d0fbb7d553deee5be5a63e4cc5f57de47ed63210

Observation 0cd33ccb-1cbf-4e55-ad77-ba6e7bff0b80 · outbound

This paper cites The learning curve: implications of a quantitative analysis,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The learning curve: implications of a quantitative analysis,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.680836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.680836Z digest=sha256:b9c6f13327019ee442084a4ed97d133a7b023b95b3d04a01c99d09dd6c85b238

Observation 27c29121-cbce-4bce-83b4-97c2df0c216f · outbound

This paper cites Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Plateaus, dips, and leaps: Where to look for inventions and discoveries during skilled performance,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.684918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.684918Z digest=sha256:0b034c9acdf7fccd91c503def95be5807dfddef764fe06a4e1c0ae1f2353f1b2

Observation 91c8e278-8ca1-4dfa-927a-b870f9af7cd7 · outbound

This paper cites A first-principles mathematical model integrates the disparate timescales of human learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A first-principles mathematical model integrates the disparate timescales of human learning,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.688862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.688862Z digest=sha256:8adfa522f2d43a3a5d5fa753fa9de4791d5aa3f528ea2a2e2d61e7fffa841174

Observation 6bd50324-3c6f-4f18-971b-9b663a5a8c9f · outbound

This paper cites Spin-glass models as error-correcting codes,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Spin-glass models as error-correcting codes,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.693324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.693324Z digest=sha256:cc0ee990d3a8be1cabd7371cb05af119df9287e3f66d248ac805e28bcf930095

Observation 8491fe46-a0cd-40d9-9638-2b9ad7f199dd · outbound

This paper cites Newell,Unified Theories of Cognition.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Newell,Unified Theories of Cognition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.697193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.697193Z digest=sha256:fa2eb9407da4c7e81ed59ad4fee6dea272bb4a897d289cf175d027cb85556de8

Observation 4c225486-1177-463a-bbeb-f745387c52fb · outbound

This paper cites Barab ´asi,Network Science.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Barab ´asi,Network Science

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.701454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.701454Z digest=sha256:3cc388f383aabe818bc490c7c62cfdf402ebc14f39d5bb9e9a739748cc1d6382

Observation b721141f-d0e6-4ec5-b9df-f670f1449ae9 · outbound

This paper cites Learning curves: Asymptotic values and rate of convergence,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Learning curves: Asymptotic values and rate of convergence,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.705522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.705522Z digest=sha256:6331184ef42813f7c0416a35db153166bc7f0eb6c593d679c3c082caab5c8fd7

Observation 27dac46d-595c-4af0-be03-16caa56f7f4b · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Deep Learning Scaling is Predictable, Empirically

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.709672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.709672Z digest=sha256:9319476534b000ab1da10eb94384df129ccbe993c92c5a86a79c0a9caeae013b

Observation 77e77b35-f651-4a00-96d2-619bda196821 · outbound

This paper cites A Constructive Prediction of the Generalization Error Across Scales.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search A Constructive Prediction of the Generalization Error Across Scales

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.713752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.713752Z digest=sha256:602b088207186fc0ba3cfe04904a5ef4051c0222adc10cc4e3edea6e1f0eff81

Observation b2483da6-1c56-4ecf-b3e0-8f16bcb95355 · outbound

This paper cites Language Models are Few-Shot Learners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models are Few-Shot Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.718254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.718254Z digest=sha256:312e2a83494aef258542ae143da310ea1c84936b3f10f29ca93595aaa8f3a85e

Observation a5a4d548-51d1-4c27-b5a3-208835a5ce7a · outbound

This paper cites Prediction and entropy of printed English,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Prediction and entropy of printed English,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.722424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.722424Z digest=sha256:3cb4cc0958dcf613f9d67d3ecc26fff6118c1bdee6662b3cc19a90286e91bde4

Observation 6366e3eb-1eff-489f-9840-95594bbd5bf6 · outbound

This paper cites Explaining neural scaling laws,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Explaining neural scaling laws,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.726859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.726859Z digest=sha256:73af5149168df96b6fbfca010f874dc4cb48096a3c3d8203468ba12eb84207a7

Observation b54a65a2-cecf-44f0-8acc-597adfd2fe59 · outbound

This paper cites Towards a universal scaling law of LLM training and inference,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Towards a universal scaling law of LLM training and inference,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.730817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.730817Z digest=sha256:39fa0996a758c1d4b9f75050d97d3ea5a848bd7abc5ba17686b0e7988457e323

Observation 1e02fe25-2a3f-4352-8ea8-8dba24901087 · outbound

This paper cites The Llama 3 Herd of Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The Llama 3 Herd of Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.734911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.734911Z digest=sha256:f3a6ec05a06f1a045069b35e4a3437880b7505e10860c552de697217f6b97168

Observation 7786d76e-84fb-4402-a578-d68742c0b14c · outbound

This paper cites Densing Law of LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Densing Law of LLMs

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.739462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.739462Z digest=sha256:c77b74fb3d35973fe8d1388589fb69fa93433154bf7b9b49952e3a3dbdea803a

Observation 341e967e-f53c-44d0-b95a-640968e74a98 · outbound

This paper cites How predictable is language model benchmark performance?.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search How predictable is language model benchmark performance?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.743953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.743953Z digest=sha256:98444841f01d6865fd9a715d26f431ea22b0901546cc7c6fd7ab72d46c6c5543

Observation 55d90686-d264-4dcc-acd3-a88ee28203c9 · outbound

This paper cites Observational Scaling Laws and the Predictability of Language Model Performance.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Observational Scaling Laws and the Predictability of Language Model Performance

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.748185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.748185Z digest=sha256:b4c9f47ee5865012b8d1eda7a86b4cc179f9fc5544226ee011fddc1a612ed40a

Observation 7212905d-590a-4677-bc73-d486d3d50ec8 · outbound

This paper cites Language models scale reliably with over-training and on downstream tasks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language models scale reliably with over-training and on downstream tasks

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.752393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.752393Z digest=sha256:fad4de71f45b3d783204bbba3c28837d8d8696bbf60f8e15a477fae2cae7d818

Observation f1e3447c-cb80-4840-93aa-5ed3ab6d4741 · outbound

This paper cites Broken Neural Scaling Laws.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Broken Neural Scaling Laws

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.756787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.756787Z digest=sha256:afd4c194d21fd8487977d1382a05123b4cede319ff4168592884fdc404322523

Observation 76fd424c-0c4b-41c9-8969-76a406233349 · outbound

This paper cites Scaling laws for downstream task performance in machine translation,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling laws for downstream task performance in machine translation,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.761555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.761555Z digest=sha256:661405d9db9dd640db3a1f40f22f03ebc2d8151d46ee57c78ba7746e52e972fc

Observation 24a0ee2a-be06-4e0c-ab17-dac540d80973 · outbound

This paper cites Exploring the Limits of Large Scale Pre-training.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Exploring the Limits of Large Scale Pre-training

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.765677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.765677Z digest=sha256:43dfe0ea91911cb95c09fefdbab8ecc37759e88f6e603eab0f95b65021e704c3

Observation 0065568d-23ea-46ab-955f-c9b2bee4b5c9 · outbound

This paper cites Scaling Laws Do Not Scale.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws Do Not Scale

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.769818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.769818Z digest=sha256:01500a86394570a2a77ad68e0ad8d119a67d772e4ed89b8c80f496d53d73a83c

Observation 663ecb2d-e843-4e8e-8586-d528424257a7 · outbound

This paper cites Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Not-just-scaling laws: Towards a better understanding of the downstream impact of language model design decisions,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.773993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.773993Z digest=sha256:0b45fbb8d926e21871043d9bae8da1f8077590ca83aafa8b0736f9329db9bf02

Observation 798b2385-8575-4476-9590-01cbf7243f93 · outbound

This paper cites Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Same Pre-training Loss, Better Downstream: Implicit Bias Matters for Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.778229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.778229Z digest=sha256:eaf5e5d5841c10ab38578a5642c044e30ca799f2740be020e6aa696afde24a43

Observation 29406675-cb2b-4863-b3dd-2c3dd95f3e92 · outbound

This paper cites Overtrained Language Models Are Harder to Fine-Tune.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Overtrained Language Models Are Harder to Fine-Tune

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.782404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.782404Z digest=sha256:b6e8fecea3d699537ef0bb5f91d2373ff9528ef11628045935f5a9ac614b0798

Observation 4638aa3f-b743-452f-b280-1e9c4233a83c · outbound

This paper cites Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Rethinking fine-tuning when scaling test-time compute: Limiting confidence improves mathematical reasoning,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.786753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.786753Z digest=sha256:f5ee4ffe82ba90533bb31f014c5a39d91306944f52fd22ec782a1276b6427aed

Observation e9dd4514-ab59-4454-848f-e22edde220ed · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.791105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.791105Z digest=sha256:ba7a55a1c4b8c82f9e8aa92627760ee6f5c43e52fb29366b9791cac28f9cef20

Observation 33d1bc75-12dc-4f7f-984d-dfbfd9f856a9 · outbound

This paper cites BBQ: A Hand-Built Bias Benchmark for Question Answering.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search BBQ: A Hand-Built Bias Benchmark for Question Answering

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.795184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.795184Z digest=sha256:582c494cebd47c2ab47a9c393ed2ff4c4e136d59e36fd6237d83928a8114fd66

Observation 50259960-870e-45ad-b591-0b717961a115 · outbound

This paper cites Inverse scaling can become U-shaped.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Inverse scaling can become U-shaped

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.799434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.799434Z digest=sha256:085d15b81bd01e88aa90ff9f5237e2589e72820fe6686e498c5cee777fb1c099

Observation b52716c6-d685-48ae-a9a6-3271f62f6f1d · outbound

This paper cites s1: Simple test-time scaling.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search s1: Simple test-time scaling

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.803725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.803725Z digest=sha256:046159e1cfdab3cfe873220ff5c6bc1e1fae04a94f7bd1adc55fd3d4b30cf325

Observation 5ad86b0f-93bd-4328-8700-4bce37b0ea1c · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.807957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.807957Z digest=sha256:c8eec5f2c5594bec157248a9490ce9ebe1f4344d2412dc17ef8393e5a3560287

Observation 5cbeb825-46fa-49b6-809b-3a8ec04067d0 · outbound

This paper cites Large Language Models are Zero-Shot Reasoners.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Large Language Models are Zero-Shot Reasoners

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.812191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.812191Z digest=sha256:a66a80b72a253ebd2cfac43891daef3d3738e02e29131ba23aff3c311b64205a

Observation 2160b047-c53e-483c-8455-5b6271a68dcb · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.815981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.815981Z digest=sha256:ecb289dbf3d2bc00da35e509381a078e51ad8f4415a12494ed8f0fc36e18bdd8

Observation 6eb5c544-c2dc-4356-ae67-c3edea4480a2 · outbound

This paper cites Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.820330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.820330Z digest=sha256:3b298b97080b818af32c1132906eece6d1661bc30905e6482bdac14cc6625a62

Observation 395dad7c-5bf2-4692-9418-3263c1e2e7c9 · outbound

This paper cites Sequence to sequence learning with neural networks: What a decade,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Sequence to sequence learning with neural networks: What a decade,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.824523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.824523Z digest=sha256:8b6f77a99e9d07692eec803353ec918176bf8912fb0d02ba8226a850d807e70a

Observation 43d03769-8801-4aad-8193-17701a25af25 · outbound

This paper cites AI doom from an LLM-plateau-ist perspective,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI doom from an LLM-plateau-ist perspective,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.828440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.828440Z digest=sha256:059c6752c93d90f6565adbb3dc68f8be88da23005a30426b6dd7fb8f64cd2add

Observation dd5603ca-4f42-4b38-a44e-be59e9426aa4 · outbound

This paper cites The first wave of AI innovation is over. here’s what comes next,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The first wave of AI innovation is over. here’s what comes next,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.832462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.832462Z digest=sha256:e091b106b30f09a97afe720615517bbd055d562d277fc1f34848da1b9cc2fba6

Observation 0a53bca0-df12-4823-90de-32112cda7561 · outbound

This paper cites AI won’t plateau — if we give it time to think,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search AI won’t plateau — if we give it time to think,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.836408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.836408Z digest=sha256:a294eb6e1e679cb7bdbf8e6668411c1d29fec093f4d273793ca051bc8709efca

Observation 22f4293a-5b65-424e-ba50-474d1949f1d5 · outbound

This paper cites Scaling data-constrained language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling data-constrained language models,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.840145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.840145Z digest=sha256:795bd82d716107ce54449000ab948816c3aede1d8c7f36c437821563e361d446

Observation 023ee840-7c37-4322-992c-200a388fecf4 · outbound

This paper cites Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Data Filtering -- Data Curation cannot be Compute Agnostic

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.844389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.844389Z digest=sha256:d6cb4ba41d92a3b6743e3d0468f077298b4b4616652a07d4cd4894174a15533d

Observation c6de41cc-bec8-40c6-bcd7-1aface91964a · outbound

This paper cites The rising costs of training frontier AI models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The rising costs of training frontier AI models

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.848494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.848494Z digest=sha256:39c28e42f33da3c1bca4f45451be463d4587feaef743c52abad2a8352a077fcb

Observation 2fbe18fe-d46e-40fe-be04-734db104bfbf · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.852434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.852434Z digest=sha256:0ca1eaae5b69b0f71dd06b9bb05241f1e2fa01670dba10a3e74d77f8ca048a11

Observation 8a60c330-fc02-4424-aa94-100361efb22b · outbound

This paper cites Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Language Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-Thought

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.856507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.856507Z digest=sha256:cbafc83196babf6bdb3e8018b31d3920ba8e7bdda90f1480175c3d0172be64a0

Observation caa6192b-6ee1-41e8-a587-77f1a904d156 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Efficient Streaming Language Models with Attention Sinks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.860710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.860710Z digest=sha256:0476b7ee74b8d005ccbc13a9b1539e44786e503a0a916c46130244ea027ce76a

Observation 71ffee8b-c47a-4cfb-a28b-3ff54ef3c5ad · outbound

This paper cites Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Don't Take Things Out of Context: Attention Intervention for Enhancing Chain-of-Thought Reasoning in Large Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.864758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.864758Z digest=sha256:912bba061c86951eee1e8aeb754583ae3d3e852a282d1dff81d937df1f6e3831

Observation 489eb03e-4f65-440f-b1aa-4fd2ce6adefb · outbound

This paper cites The curse of CoT: On the limitations of chain-of-thought in in-context learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search The curse of CoT: On the limitations of chain-of-thought in in-context learning,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.868843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.868843Z digest=sha256:9a1124e5bc84bfc48e50c0dfa4a5caaf2be7dffe07beded834352c4bf38df484

Observation 634eb7db-c73e-4d8a-92e9-9d8456bc662d · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.872688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.872688Z digest=sha256:91285d8e0a9b3984fdb022afed3b3f084744a444709c53e8a100c251ada61ca1

Observation 79cc6bdc-7dea-481c-bbe1-42664f1d2e42 · outbound

This paper cites Let's Verify Step by Step.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Let's Verify Step by Step

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.876992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.876992Z digest=sha256:0ca0b4f074bca7cfbd6aecb2c05112554a56050798da4bf7a4bbb72d70e914c8

Observation a884701b-03ed-4fe4-9a68-c650ccb5ee85 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Training Verifiers to Solve Math Word Problems

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.880977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.880977Z digest=sha256:31693d6a9ae2ebcd9c582af49613b5a673856198acbb56fc00c3764966669cbc

Observation e262f215-9b8c-4ac1-88db-db32dce362a4 · outbound

This paper cites When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search When to solve, when to verify: Compute-optimal problem solving and generative verification for LLM reasoning,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.884774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.884774Z digest=sha256:3522ae2605f6817b505ab8fffc60b03e4e068588e136bdd83ca67c250c196ac4

Observation f052279f-279d-41c7-83a1-2efabc1fa6e6 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.888499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.888499Z digest=sha256:79915539440b889bfb376e941af2a8d131a2c1406b5bc8470e716f57927f931a

Observation bf2247ff-05b5-4065-9db4-a24a122bd5a1 · outbound

This paper cites Solving quantitative reasoning problems with language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Solving quantitative reasoning problems with language models,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.892256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.892256Z digest=sha256:439313ff8a5d9476e6f37d8fd111b80d8472d642a7b48ed289f0391e6acc1515

Observation 055e7f58-e5b6-429d-bf48-6546b15f8b3d · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Self-Refine: Iterative Refinement with Self-Feedback

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.895889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.895889Z digest=sha256:d238196014c42f0f64b8441d4322ffe19defb568dae054b982a8a6bbb916658e

Observation bf4a5fc3-212a-4143-b14f-e9186f1fb833 · outbound

This paper cites Cost-of-Pass: An economic framework for evaluating language models,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Cost-of-Pass: An economic framework for evaluating language models,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.899967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.899967Z digest=sha256:597dd6317a11d32bb9f67f67531dc0cf9858c611e17348292f5af0ac90c7de88

Observation 2aa08e6c-7e5a-4790-be4b-c704137801f3 · outbound

This paper cites Scaling Laws for Transfer.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Scaling Laws for Transfer

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.903603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.903603Z digest=sha256:dda42466f16fcdcb7eda853a52a2b963ab3d3f32b72a61c0e50587d05c1d5560

Observation 45d542b2-0c07-4f17-ba0f-eb4f23d3ebc4 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning,.

A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search Reproducible scaling laws for contrastive language-image learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:39.908059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:07:39.908059Z digest=sha256:5ff93aec7facbf1e9226992a32643b9f85d2728d909c293c863071fe22648d88

Pith citing papers

Observation 5e5e584d-abf1-49fc-9722-19276522f94e · inbound

Two AI Metrics Diverged: Will it Make All the Difference? cites this paper.

Two AI Metrics Diverged: Will it Make All the Difference? A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:36:56.137222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-07-02T12:29:24.439779Z digest=sha256:ca831d53d338451ff4937239e00734c772da935ee3b1cc19bbb38d7d71cb19d0