Pith. sign in

Paper Citation Record · LEDGER

LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2410.02884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02884 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:35:28.447005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:29:57.287817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c196a863-f5ca-4cad-813a-6f1ed1e12eb4 · inbound

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems cites this paper.

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:35:31.446927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:35:31.375020Z digest=sha256:7139fad0d3d70a18fcd883a9ff78eac3ec27ce466812f9840308a14bd02b07f8

Observation bc8882e3-3828-4ee9-a73b-c4d5ca976e0a · inbound

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs cites this paper.

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:36:50.157350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T12:36:50.060335Z digest=sha256:a56a501b27bf6d2461294108250955425d885edc5aa87e4b7d8e8235001ca293

Observation ea5fe782-2f0d-4b24-ba71-c408cb908a1d · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.643883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:c292a3f6349374b85e392986b499ce78b274bc8275aae9e902e086b7c8c283ea

Observation 1cd4adb3-21b7-4b9a-82ee-e6d3d2c29b27 · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 187

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.427455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:1275405f59235ce40a7715867ec9b6549bae49fb9bfc59dceba3d7495e448d00

Observation 4bfc6eeb-f2ca-4590-b52d-89aa034c0b8f · inbound

Large Language Model-Enhanced Multi-Armed Bandits cites this paper.

Large Language Model-Enhanced Multi-Armed Bandits LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-09T16:35:28.447005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:35:28.447005Z digest=sha256:26207b7af26e4688218625da2f3548a32bd99149ada37c99abf7c330499439c1

Observation cf00df94-b20b-4b68-8a09-efa45ca0b981 · inbound

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search cites this paper.

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T11:57:48.815737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:57:48.815737Z digest=sha256:ba012984a6063d8cd02af304752d4f9ebfe0326a350bffbaef8b09dd0f85c319

Observation b6ff821c-16ab-4b30-9cdc-9bb1c1ebb1ff · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 246

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.542825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:ce159b4799d67209b05b395a234bdacf735397a248f0b51513fc7ed1c3bf99e2

Observation d8588eec-ff8a-4a69-9abf-5cc2cfbe793b · inbound

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools cites this paper.

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T22:03:46.492505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:03:46.492505Z digest=sha256:a4f033cc7e6130d03ee4a492690c6b099642208b00dc3bc4cc50026b7094b94a

Observation 170a4518-75c1-4145-ac98-90b41a6f8266 · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.985271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.985271Z digest=sha256:e486a3ea94803e048e94c2a8c3b4a9bdd767ef7aebaec32b247703b7cbd6c059

Observation 407f26e5-679b-458a-860b-e946dfad7d99 · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.631264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.631264Z digest=sha256:6bcab896dcfb22a60b6205b74319a025c346dd698c55d1a30d8994b6ab224286

Observation cb1b0403-8a74-4fa9-bec1-c619e07898e1 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 232

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.378838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:5d6bb5ace8e89e552c207a1b6e3d722b184f76302c1628c612f7a894d7030123

Observation af48fcaf-14c9-4ccc-b398-0d066deade89 · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.313394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:f54e02a73caf8b194f6b8f567da0dda1f3591a4edb1058f448071c94dcab6d76

Observation bd1a5786-fbb5-42d5-b9c1-36175921f33d · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:59:03.329513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:633ef654d7925c46d6676f131a02f882b35fb877ab71575f5325c3f040d1644d

Observation bc6b0951-13c5-416e-a26c-64c85a88fa62 · inbound

Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems cites this paper.

Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:42:10.437330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T21:39:49.832151Z digest=sha256:3dedaa1c7fe7dadd474041302cff390cc39389a686e8c4a370aa301d829b1273

Observation 3dc80ecb-14b2-40a2-a3be-0fd6924aef7c · inbound

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning cites this paper.

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:02.552762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:02.552762Z digest=sha256:4b54b5c3903f5c1c47f3e8a02f6e5d9e13c0b3fea5bbadcbf087f02f2ac7dc77

Observation b23991e7-dce8-4463-9a82-445e0a635091 · inbound

Reward Model Generalization for Compute-Aware Test-Time Reasoning cites this paper.

Reward Model Generalization for Compute-Aware Test-Time Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:27.539519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:27.539519Z digest=sha256:bf74693fcca3727b68388573c745ace7c9cdafc0d9f09d1d878d2e5c20e98971

Observation e10c0273-71c9-40c8-b89f-07d805541b8d · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:36.471077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:36.471077Z digest=sha256:6637ab75b7a8eba557f8bd25fc76abb50a1623556d1767fb3b72199fca276fe0

Observation 25b11409-beeb-43db-9ca3-f480f5916501 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:58.712192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:58.712192Z digest=sha256:59c9906eb1c4144fece32cbc0df7626eb464b10415883e382a808082b3674c94

Observation 9ebaadc9-87c8-41a0-ba69-f23f61aa0c81 · inbound

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM cites this paper.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.857251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.857251Z digest=sha256:b9b8b40146c37d989f47c240fb176b702551ad456b338389f07c2b45de3cb608

Observation ab44c3fb-adf0-49fd-91b6-4a4f2c95a478 · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.401314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.401314Z digest=sha256:2e8b64247b4fe0a37e77896592e53e540da252a45bf677dbd9eab7889bb51dbf

Observation f8eaf83f-44f6-4818-8360-306b7170690b · inbound

SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition cites this paper.

SELT: Self-Evaluation Tree Search for LLMs with Task Decomposition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:38:34.047050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:38:34.047050Z digest=sha256:9b43bf10bb7046e63210989950b96fe528e91384a38e93efc4bf5ab8a927ab54

Observation 898d3920-df61-443f-a8bb-61b92375818a · inbound

A Survey on Large Language Models for Mathematical Reasoning cites this paper.

A Survey on Large Language Models for Mathematical Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:47.574551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:14:47.574551Z digest=sha256:30a28a080bc9cd2a9babe04773e38c7c709f037cee1c12f8494b692ce7e022e8

Observation f87f5801-3e66-45b7-a293-8186d484b3df · inbound

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism cites this paper.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:23.349125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:23.349125Z digest=sha256:32df447717b8f1f0b962d343aba9e8804d60d012cad8525c4242588f964975e4

Observation e9c90b97-f75b-49c0-b883-4ae79b60d37e · inbound

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning cites this paper.

Boosting LLM's Molecular Structure Elucidation with Knowledge Enhanced Tree Search Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:59:13.738456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:59:13.738456Z digest=sha256:238d522b8767279fbd5d1e297b8bb8c9a967b383d3d25ca6bf6423dfcbafbf00

Observation ffeeb267-0a6e-4780-8502-2ca705872d90 · inbound

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality cites this paper.

Building Task Bots with Self-learning for Enhanced Adaptability, Extensibility, and Factuality LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 207

Resolution
unresolved
no resolver link, observed 2026-08-05T15:38:55.248368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:38:55.248368Z digest=sha256:bad0a853f6d53f1deef5fe1381eab6b0d3311991f98f854df51b913df9861926

Observation f2bd868a-d220-45f5-ab20-d92fb7837cc5 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:00.088338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:00.088338Z digest=sha256:02c4af2b0485626d544b8cff1d2831fe4d014be10cbe6b9b2a97c38668a0e142

Observation b4439750-3acd-4337-bf61-23ea4695e8ba · inbound

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning cites this paper.

Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T12:52:28.323140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:52:28.323140Z digest=sha256:88ed76391236e13dcec0348b570a3b463e1ce93fe7c8f8947088eb6e07814330

Observation 8602e035-1add-4b8e-b107-eb26c08f2193 · inbound

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance cites this paper.

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T10:53:10.440907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:53:10.440907Z digest=sha256:4ca36a12767bfd7cf31b4cd3cb9c6f273e35e12ac5229709961b6e23b0bef0a0

Observation d58a5c61-c0a5-4142-9673-253d1e27afb8 · inbound

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning cites this paper.

SIGMA: Search-Augmented On-Demand Knowledge Integration for Agentic Mathematical Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T06:57:16.504758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:57:16.504758Z digest=sha256:70a67971cbf480437b2757a48cd6913acf5eb0de51cad6b0227f0b4fc9ce0653

Observation 886e57de-74cb-493d-b40b-984eca64f435 · inbound

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators cites this paper.

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:57.590779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T02:05:48.483640Z digest=sha256:d218ec752dde20aeb407c924f94db6d4b69f43fa493942ceeadeca1a56228c7f

Observation 13c9f9ad-3931-456f-ae39-c1d6ec93214e · inbound

Latent Visual States for Efficient Multimodal Reasoning cites this paper.

Latent Visual States for Efficient Multimodal Reasoning LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.289397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:38:11.619574Z digest=sha256:d8b4e8391d45efcf5903e91d7567b277617d12c30f89ca414fac3c88494efc19