Pith. sign in

Paper Citation Record · LEDGER

Learning Adaptive Parallel Reasoning with Language Models

As of 16 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 17 inbound Pith citation observations for arXiv:2504.15466.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15466 v2

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:31:32.271792Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:40:10.420547Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:29:38.199469Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 68309b73-aabc-44c3-a7a7-56ff8b667800 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Learning Adaptive Parallel Reasoning with Language Models Graph of thoughts: Solving elaborate problems with large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:33.047257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.098276Z digest=sha256:07285a2b36ccd98ebf32a2bd01ae039235bfb275b9c6f89275ab1c881e7b6ba9

Observation 55424902-cf9e-422f-aee5-a7f19bf123a6 · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling.

Learning Adaptive Parallel Reasoning with Language Models Decision transformer: Reinforcement learning via sequence modeling

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:33.033586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.103333Z digest=sha256:deb5f59df1974e148946cd081f729ad81ddbe868e182abf19ace6759ee2726e9

Observation 84fc970e-49a3-4668-8566-c2c66c36a7a0 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Learning Adaptive Parallel Reasoning with Language Models Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.108037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.108037Z digest=sha256:50b78eea40cedb753b061928f48c5f3bb027ac6f16cdde0a7a36c455be7c86c5

Observation fc3d7509-fe5a-4a8a-9f1c-f45c6a2e0692 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Learning Adaptive Parallel Reasoning with Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.113005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.113005Z digest=sha256:f8685bd5b4a3c3b1bc4471e01da0f46dccacae0040b793a30e43fb6f497787f4

Observation 5faccf53-ae96-46d4-a5c3-21bba3cd9419 · outbound

This paper cites Torralba, J.

Learning Adaptive Parallel Reasoning with Language Models Torralba, J

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:33.017737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.117767Z digest=sha256:6e0c179144322bdf117818ed0ad4336583c22398649acbb92da984b6a828a77a

Observation fc06bab4-7461-4493-9479-b104edd10e80 · outbound

This paper cites Stream of S earch ( SoS ): Learning to search in language.

Learning Adaptive Parallel Reasoning with Language Models Stream of S earch ( SoS ): Learning to search in language

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:33.004218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.122092Z digest=sha256:5a5e6e9748ce5b34ff10108327e491e5dc651151f1b60bef2af9b1305faca14f

Observation 4bde2667-9fa3-4fc0-8e9e-0e83f3763999 · outbound

This paper cites Think before you speak: Training language models with pause tokens.

Learning Adaptive Parallel Reasoning with Language Models Think before you speak: Training language models with pause tokens

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.990067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.126583Z digest=sha256:ae004e93cc699f5ee938d81bc4b7552360abfee151db8af37e063953c1405894

Observation 7b10acde-ec64-47ab-a9b6-ec85012eb87c · outbound

This paper cites Self-Steering Language Models.

Learning Adaptive Parallel Reasoning with Language Models Self-Steering Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.131620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.131620Z digest=sha256:4bbb9c37ef8be7765a2a01606179e2f5d67da86496d84c7d9f84bfd20220da2e

Observation 82b0a73c-9f19-4949-8384-9b43749c7f85 · outbound

This paper cites ETS: Efficient Tree Search for Inference-Time Scaling.

Learning Adaptive Parallel Reasoning with Language Models ETS: Efficient Tree Search for Inference-Time Scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.136274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.136274Z digest=sha256:b8f677237906d8c17e4e2eba009c957434e01fd0c7247f32088886caed50e7f8

Observation a890ab5c-de1c-4106-b7c0-de048c840424 · outbound

This paper cites Interactive Speculative Planning: Enhance Agent Efficiency through Co-design of System and User Interface.

Learning Adaptive Parallel Reasoning with Language Models Interactive Speculative Planning: Enhance Agent Efficiency through Co-design of System and User Interface

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.140642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.140642Z digest=sha256:796aa5f39881bfa3fd89d1e8f1e96721211a3c0471e36175803d9bccda4bbd2d

Observation 56e8f61c-41db-4795-9e7b-e237d57fd58d · outbound

This paper cites Learning to Keep a Promise: Scaling Language Model Decoding Parallelism with Learned Asynchronous Decoding.

Learning Adaptive Parallel Reasoning with Language Models Learning to Keep a Promise: Scaling Language Model Decoding Parallelism with Learned Asynchronous Decoding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.145239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.145239Z digest=sha256:b8cea7ba47947e53deeb60d3e55b9cdfac3d7ca2f6f2a11e5e766c95a8585b61

Observation 131d966d-05cf-44c0-9a3f-9b1910054207 · outbound

This paper cites Mahoney, Kurt Keutzer, and Amir Gholami.

Learning Adaptive Parallel Reasoning with Language Models Mahoney, Kurt Keutzer, and Amir Gholami

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.974809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.150126Z digest=sha256:6855572af061b020ebb745cd6daec8ccc8e9fb36e54ca8c9267412433aab1b4e

Observation 1a142b8c-8074-4d01-94c9-975af057eeb3 · outbound

This paper cites Efficient memory management for large language model serving with PagedAttention.

Learning Adaptive Parallel Reasoning with Language Models Efficient memory management for large language model serving with PagedAttention

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.960505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.154420Z digest=sha256:2d3e0977d6cc1f9e58f16e74e039b854c4cb85f3f8bb28200479ecea7c863e2d

Observation acd64d26-39fd-4929-a583-65035de0b42e · outbound

This paper cites S*: Test Time Scaling for Code Generation.

Learning Adaptive Parallel Reasoning with Language Models S*: Test Time Scaling for Code Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.158683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.158683Z digest=sha256:930daa1e83fdb10f1060d4c493579269f090dcedeba5b730122b5d228774d6ec

Observation 59ff437c-7b38-400b-8206-8c64a89518bd · outbound

This paper cites Skeleton-of-thought: Prompting LLM s for efficient parallel generation.

Learning Adaptive Parallel Reasoning with Language Models Skeleton-of-thought: Prompting LLM s for efficient parallel generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.946249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.163014Z digest=sha256:97d20f9cd5cb28ad8bc2fdb5c321eded36eb905a21dae326fa7df9c2b3159440

Observation cacf41a7-a8eb-4a14-bf1e-a61142623dfc · outbound

This paper cites Learning to reason with LLMs , 2024.

Learning Adaptive Parallel Reasoning with Language Models Learning to reason with LLMs , 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.931478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.167449Z digest=sha256:ae98442f71913bc18e194724c989560fcb73e4588892064afebe70d6b9322bf1

Observation 694d11c5-7d83-4754-b4d8-6967dd0d3f43 · outbound

This paper cites TinyZero , 2025.

Learning Adaptive Parallel Reasoning with Language Models TinyZero , 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.916505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.171542Z digest=sha256:0c4037f7542d9a74fe22a83401a2aca5f1701f07d5328041f3dfb3bf4bc3348d

Observation 6e2cd077-9f06-4fa0-b1e5-bab45879cb63 · outbound

This paper cites Measuring and narrowing the compositionality gap in language models.

Learning Adaptive Parallel Reasoning with Language Models Measuring and narrowing the compositionality gap in language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.900606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.175741Z digest=sha256:f52fce44f8aba38b8cd5692ecdebb81961aeb87a7badc9b12f5244fb1424181f

Observation 45e27e38-8f0b-47f3-a7cd-57f78ff0b601 · outbound

This paper cites Hogwild! I nference: Parallel LLM generation via concurrent attention.

Learning Adaptive Parallel Reasoning with Language Models Hogwild! I nference: Parallel LLM generation via concurrent attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.180016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.180016Z digest=sha256:7344d127bca4634b32383477acb8021dbd270f3486b8988ab4f817bf01d6477d

Observation a1dc24e5-68a7-49f1-83ce-a14bb5bdc6c4 · outbound

This paper cites THREAD: Thinking Deeper with Recursive Spawning.

Learning Adaptive Parallel Reasoning with Language Models THREAD: Thinking Deeper with Recursive Spawning

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:31:32.511658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.184296Z digest=sha256:ae639ba0485d1e5243202486e6ad74f1263af03ea6927dd94cf7927d779d386f

Observation 4fe769f0-97b3-4843-867a-22214f603214 · outbound

This paper cites Algorithm of thoughts: Enhancing exploration of ideas in large language models.

Learning Adaptive Parallel Reasoning with Language Models Algorithm of thoughts: Enhancing exploration of ideas in large language models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.884153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.188917Z digest=sha256:440f941a97ab515a058e1e8bf043322b9fb8ca704a0f2570d1c8e19ebc168c90

Observation 0889afa5-f94c-4fd6-bce6-3aca65458599 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Learning Adaptive Parallel Reasoning with Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.193018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.193018Z digest=sha256:7b0e67afcc0f911354928ac53b7dd9934d555dfadc9209d235c43300abcb8b01

Observation 0c00b2a6-d2e5-4e1c-a899-8868eb1cd398 · outbound

This paper cites Scaling LLM test-time compute optimally can be more effective than scaling model parameters.

Learning Adaptive Parallel Reasoning with Language Models Scaling LLM test-time compute optimally can be more effective than scaling model parameters

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.870104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.197481Z digest=sha256:6e6a839b44a076c1f9726bb70a18f6320694ad32ca5c77ca854fafeaf839a63d

Observation d9884ece-8e85-458d-8cfa-807093b84cd2 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Learning Adaptive Parallel Reasoning with Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.201746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.201746Z digest=sha256:947dfcfcc02508261b65609f4ce7b4039a82123ee05822f4ca6560d76bc8a9c1

Observation 32ae8dd2-6771-40ae-9a4e-ee7e8809dd70 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Learning Adaptive Parallel Reasoning with Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.206169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.206169Z digest=sha256:b884c305be312b342144f2ebc7ef2a3f046281e4c445cec5fba1cb341881ba64

Observation 5fbd31f7-73d7-49fb-b749-e7dff4936a67 · outbound

This paper cites Atom of thoughts for Markov LLM test-time scaling.

Learning Adaptive Parallel Reasoning with Language Models Atom of thoughts for Markov LLM test-time scaling

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.210680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.210680Z digest=sha256:242b5397d5bd1cf482c08ddb81cf1fbb625f96a0a69e855ba53820e3b0e3f877

Observation 40983bde-72e2-457e-ab5c-4602c0a04b14 · outbound

This paper cites Mixture-of-agents enhances large language model capabilities.

Learning Adaptive Parallel Reasoning with Language Models Mixture-of-agents enhances large language model capabilities

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.855770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.214806Z digest=sha256:81e6b0afdb626539c9e81a339ff654c00f6abfc2c54698cbb220a95a630fea7a

Observation ae9e875d-26b1-416f-83f7-97404eb048f8 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Learning Adaptive Parallel Reasoning with Language Models Self-consistency improves chain of thought reasoning in language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.841657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.220098Z digest=sha256:c2b122dc7aa4ce0e40ceb12f5260a880dfb2233424c73cb1882e7e5ae33d474b

Observation 82000e59-cfe7-4af4-b015-0bfcbacf85d2 · outbound

This paper cites Chi, Quoc V.

Learning Adaptive Parallel Reasoning with Language Models Chi, Quoc V

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.223950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.223950Z digest=sha256:aa4b37e4e9e13e061c01be1b15bcbd892d7b8877059ea100ac8210d17b6864b3

Observation e1bebec1-b1c0-4d64-93b8-2c69ff64725c · outbound

This paper cites PENCIL : Long thoughts with short memory.

Learning Adaptive Parallel Reasoning with Language Models PENCIL : Long thoughts with short memory

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.816771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.228058Z digest=sha256:7a7703603029d43e03f54aca453d0cc21bc6f2253d3c5344481a38597fdfb915

Observation ec56c1a0-0a54-445d-ad72-5c50b0df46b3 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Learning Adaptive Parallel Reasoning with Language Models Tree of thoughts: Deliberate problem solving with large language models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.801005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.232059Z digest=sha256:e43853584c086eb424a1ba2fd690f5639039e4f2ab3eda7a628f500adffc0de6

Observation bcc6f111-12d7-4d46-ac20-da8afeea44b5 · outbound

This paper cites ST ar: Bootstrapping reasoning with reasoning.

Learning Adaptive Parallel Reasoning with Language Models ST ar: Bootstrapping reasoning with reasoning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.785190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.236301Z digest=sha256:3d71f7c5e792f2facb3fc8a35c49a088e8f069e70490ae77d842a4c35e986892

Observation 08361a7d-f05d-4c82-aaaf-8d5d379ccc9a · outbound

This paper cites Chain of agents: Large language models collaborating on long-context tasks.

Learning Adaptive Parallel Reasoning with Language Models Chain of agents: Large language models collaborating on long-context tasks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.770374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.240554Z digest=sha256:ddd6501a2b04b578ce80b0eb88dd4ebb6b1579fd22151c927eaa0c3a823de410

Observation 299a6dcb-97a1-48e0-ab16-4d54c046f5e0 · outbound

This paper cites Gonzalez, Clark Barrett, and Ying Sheng.

Learning Adaptive Parallel Reasoning with Language Models Gonzalez, Clark Barrett, and Ying Sheng

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.755010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.244540Z digest=sha256:494aa24e854b7b01e118abf63cfe67f7175a8aed2fd8f79a8a6ce9f1703b2286

Observation d107038c-6968-4e10-a369-17acf450b163 · outbound

This paper cites Fine-tuning language models with advantage-induced policy alignment, 2024.

Learning Adaptive Parallel Reasoning with Language Models Fine-tuning language models with advantage-induced policy alignment, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.739363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.249009Z digest=sha256:2253c657ed512e20a1bbcab8cd630bfcab64ae0b9307f894a5763f339aa40fe0

Observation ba6fff12-344b-4e13-98a4-ec45b5405954 · outbound

This paper cites Language agents as optimizable graphs.

Learning Adaptive Parallel Reasoning with Language Models Language agents as optimizable graphs

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:31:32.724370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-16T11:31:32.253538Z digest=sha256:01d61f06bf3f147f52234c3975dfb12e0bd25a7ad0f954553997426657afb617

Observation 125579db-fe19-4932-9bfa-d92264147096 · outbound

This paper cites write newline.

Learning Adaptive Parallel Reasoning with Language Models write newline

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.257705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.257705Z digest=sha256:48abbf5a6fb4bfc0e1a55fe9c4ff72d6a881f9ed7f1591a41f81d3b262f9b08f

Observation 3864634c-7fa0-4cc6-9342-caa4d727b980 · outbound

This paper cites @esa (Ref.

Learning Adaptive Parallel Reasoning with Language Models @esa (Ref

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.262912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.262912Z digest=sha256:62ff1e358fd65ae6b2354944bad63bc6d88618812f8a55205b0066eccdfb34b7

Observation c9531081-715c-4d5e-ab3f-5bcfd4bbdc4c · outbound

This paper cites an unresolved cited work.

Learning Adaptive Parallel Reasoning with Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.267533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.267533Z digest=sha256:d53a581a4f8701e3b4316090fdb326c86ff53738437f19459dbcf60b6a1f9784

Observation b3c5ee7b-4f10-4499-9976-b4056470c9eb · outbound

This paper cites Goal reached.

Learning Adaptive Parallel Reasoning with Language Models Goal reached

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:31:32.271792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:31:32.271792Z digest=sha256:091320073d3a612a6ad8dff9a419138683f26d2788c74524e8f969390cab238c

Pith citing papers

Observation 778072e3-ba3b-4f34-ae77-2515f68236b8 · inbound

VeriThinker: Learning to Verify Makes Reasoning Model Efficient cites this paper.

VeriThinker: Learning to Verify Makes Reasoning Model Efficient Learning Adaptive Parallel Reasoning with Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:13.809795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:13.809795Z digest=sha256:52f5463f8c7280855ad945f3c0a5c968787f6faa9b286fe7a4d4f6764d81df60

Observation d2a83943-0b4d-4d82-8199-f60b67b5ea85 · inbound

Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models cites this paper.

Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models Learning Adaptive Parallel Reasoning with Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:36.389503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:58:36.389503Z digest=sha256:f6f926941583da24c4111cdf70e29de5df0f935a06ef6ceda45f10f732fc8f4a

Observation aa751dfe-6a10-4d89-a3c1-a93ec7d09798 · inbound

Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning cites this paper.

Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning Learning Adaptive Parallel Reasoning with Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:38.172948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:38.172948Z digest=sha256:fa60eade0b3594314333255ce0bfa823377b865f1bbdeb2382920851517bf781

Observation 087a7916-5195-44d3-b404-6393ec5f32f2 · inbound

Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs cites this paper.

Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs Learning Adaptive Parallel Reasoning with Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:10.612245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:10.612245Z digest=sha256:1a5fabb8c664bc4f4b209456ae0e7de9dab72a6f621f4f0b7d83d53a139d9d16

Observation e237c6bd-3754-4f22-8b47-4d34479a2bfe · inbound

Adaptive Termination for Multi-round Parallel Reasoning: An Universal Semantic Entropy-Guided Framework cites this paper.

Adaptive Termination for Multi-round Parallel Reasoning: An Universal Semantic Entropy-Guided Framework Learning Adaptive Parallel Reasoning with Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:59:42.167307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:59:42.167307Z digest=sha256:4aea6c3598da4021c0de69379940c5eb6e7f10246758dfff4e16c7ae162c8188

Observation c9001b59-ee10-413b-ac5c-d623503dfb07 · inbound

ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute cites this paper.

ParaThinker: Native Parallel Thinking as a New Paradigm to Scale LLM Test-time Compute Learning Adaptive Parallel Reasoning with Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T13:52:05.548609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:52:05.548609Z digest=sha256:90334d27cb8ff01bac011d993b4ebbb6b01120f97c9e20f04ba9637e1aa50024

Observation 6fc2d1c6-c9bc-49c0-ace7-5ab8591c63f1 · inbound

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning cites this paper.

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning Learning Adaptive Parallel Reasoning with Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T21:28:49.523071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:28:49.523071Z digest=sha256:6ac18243f0c43915a9120110f58c2bb17950090cbe666f45afad1ee614b0a5a0

Observation 69bf5826-4b97-4793-b273-3b5805e18be1 · inbound

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts cites this paper.

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts Learning Adaptive Parallel Reasoning with Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:42:38.729068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T13:42:07.883909Z digest=sha256:826871f0998c8b2cdf04b1353ac42cc631cd05b61881ee6a7955ce174b7d9b0b

Observation 44b50064-72a4-4f0f-8d65-5a3587cc2fbc · inbound

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning cites this paper.

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning Learning Adaptive Parallel Reasoning with Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:13:47.963496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-17T01:11:26.411893Z digest=sha256:deef7149bdcbcaf2ff5bd4ae750b9b5f55656e9878464d356fc2d3ac139a95f6

Observation 2bb5e1b6-5395-473f-8513-d23079c4eccd · inbound

Efficient Reasoning on the Edge cites this paper.

Efficient Reasoning on the Edge Learning Adaptive Parallel Reasoning with Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-13T23:28:12.790404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:28:12.790404Z digest=sha256:5d3325809eb53e5cab7ea2e39e0b9af9b9c103d7b601aba6a616235f63874925

Observation ee5b8261-6083-4f0b-8ae6-e4ba82392999 · inbound

Test-time Scaling over Perception: Resolving the Grounding Paradox in Thinking with Images cites this paper.

Test-time Scaling over Perception: Resolving the Grounding Paradox in Thinking with Images Learning Adaptive Parallel Reasoning with Language Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:26:01.511887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T16:38:11.785469Z digest=sha256:ed6a0a2375e6c6e47d2bc4a1813485816db5e4c0246db66c7801261f1f30cb1e

Observation 24243ed3-71e0-458c-b697-36d0eb666a2c · inbound

Test-time Scaling over Perception: Resolving the Grounding Paradox in Thinking with Images cites this paper.

Test-time Scaling over Perception: Resolving the Grounding Paradox in Thinking with Images Learning Adaptive Parallel Reasoning with Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T05:31:43.163303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:31:43.163303Z digest=sha256:8da26d3ee66051ce8d00d6d0fdb558215c25e834a97b4aca8c0cc0cab95bc551

Observation d6ce8bac-a13f-4632-9f2d-16f6d604241c · inbound

Regulating Branch Parallelism in LLM Serving cites this paper.

Regulating Branch Parallelism in LLM Serving Learning Adaptive Parallel Reasoning with Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:58.356445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-11T01:00:53.308946Z digest=sha256:c3091d6811fab19245cd43ecc70e0b707c97a994e049091339e2be13386c6fab

Observation f2cfce90-932c-4a10-b0c4-b9e8114656d6 · inbound

CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning cites this paper.

CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning Learning Adaptive Parallel Reasoning with Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T15:42:38.398701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-19T15:39:56.255871Z digest=sha256:87b635bb1c4e687623972932593e78b7cbb990bbb095aa7ebb1609549111f03f

Observation e4b88ca3-5565-46ed-a1c1-f3b1252cc84e · inbound

Sakana Fugu Technical Report cites this paper.

Sakana Fugu Technical Report Learning Adaptive Parallel Reasoning with Language Models

Reference 264

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:29:38.201139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-26T14:22:37.596720Z digest=sha256:cc1e613460a3a77fc151ef29ef93c0fd0d8e5f9fec09694eab4a8b81f576847b

Observation 49661fef-232b-4b40-acd3-822c432c5054 · inbound

ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs cites this paper.

ParVL: Parallel Scaling and Expandable Compute Allocation for Multimodal LLMs Learning Adaptive Parallel Reasoning with Language Models

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-05T04:16:08.562607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:16:08.562607Z digest=sha256:c62250008e1fc295f471f87adc27bcdf0b4b3cbdf5a74e37e784856630494ae9

Observation 2866dac1-beb5-42c4-8b34-fb80d8951f81 · inbound

Hidden Language Consistency Phenomena in Reasoning LLMs cites this paper.

Hidden Language Consistency Phenomena in Reasoning LLMs Learning Adaptive Parallel Reasoning with Language Models

Reference 286

Resolution
unresolved
no resolver link, observed 2026-08-14T04:40:10.420547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:40:10.420547Z digest=sha256:e548e11ee5408f7b20576fb18f125bec1bb13402bb5d87d92ad5fb6e3702fa4b