Pith. sign in

Paper Citation Record · LEDGER

O1 Replication Journey: A Strategic Progress Report -- Part 1

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 59 inbound Pith citation observations for arXiv:2410.18982.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.18982 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 59 of 59 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:43.679640Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T11:09:46.422756Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 24a0e5eb-5baf-4302-a3ac-26546fb9a56d · inbound

Do LLMs Really Think Step-by-step In Implicit Reasoning? cites this paper.

Do LLMs Really Think Step-by-step In Implicit Reasoning? O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T13:52:37.946531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:52:37.946531Z digest=sha256:96eb2ef9fea85bd9e916286a682206cff245e026ca0ac8588fe539cb0c979547

Observation 07096a4b-65a8-441f-8c82-c247288a4379 · inbound

BlendServe: Optimizing Offline Inference for Auto-regressive Large Models with Resource-aware Batching cites this paper.

BlendServe: Optimizing Offline Inference for Auto-regressive Large Models with Resource-aware Batching O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T13:36:44.038592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T13:36:44.038592Z digest=sha256:b7926ce4e9c08da2206de8b00d93aedfe9961c055a3f11165af00765a502d4a4

Observation f744c515-47f7-455d-a333-d908b74df1e8 · inbound

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? cites this paper.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.860045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.860045Z digest=sha256:d11f8f7534665a036696f8118b4d568ab334277449c510399f40243c97121788

Observation e2296d20-c191-423e-a7db-050e5dd5f548 · inbound

Don't Command, Cultivate: An Exploratory Study of System-2 Alignment cites this paper.

Don't Command, Cultivate: An Exploratory Study of System-2 Alignment O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T12:38:19.262867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:38:19.262867Z digest=sha256:65c253953a76cd47f71c3742078ddac59788a26226faa5ef2c7088c895a59d78

Observation af52680b-0109-42a1-b44d-0d14b27359c0 · inbound

Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS cites this paper.

Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:13:42.367173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:13:42.367173Z digest=sha256:3c619d3ca77af96d16fa9fc3c7c17f403861c93a09023714c015983faabaaea9

Observation fbc91d87-fdc7-4ae0-96ea-3457ee901e3e · inbound

Process-Supervised Reward Models for Verifying Clinical Note Generation: A Scalable Approach Guided by Domain Expertise cites this paper.

Process-Supervised Reward Models for Verifying Clinical Note Generation: A Scalable Approach Guided by Domain Expertise O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T14:03:18.885861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:03:18.885861Z digest=sha256:8264fdde4e980a8c2ede3723460ff7fc1615a684a7468ec5ec70f1a9e6497246

Observation 2bd6778b-26b4-4a5b-8dd4-f8ae8b2183dc · inbound

Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation cites this paper.

Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T11:40:13.712702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:40:13.712702Z digest=sha256:73b5a7baf9cae3a2d7df9e6af535a894b53edc858dcbcfd75e7d62ca3f3e512d

Observation a2e42d08-de1c-4709-938a-a9d07cce0d1a · inbound

DRT: Deep Reasoning Translation via Long Chain-of-Thought cites this paper.

DRT: Deep Reasoning Translation via Long Chain-of-Thought O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T05:32:00.370371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:32:00.370371Z digest=sha256:13145be765d7fff2be9925dad4d286c6d710334a5b5ada21168e770a7cc72050

Observation 8949b729-f708-4a48-918d-06351c144546 · inbound

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs cites this paper.

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T12:36:50.238428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T12:36:50.060335Z digest=sha256:b1fda9d31ca86926bbbadc0ee9d9393d0f242043b58dbf32b7b0e47260313d6b

Observation a2c26d62-487f-4dbd-a5db-97a4b8e51cce · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:36:27.578579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:dad07b895b79abfef6def837c502649080db5ca37295843030c17b7c00a1b60b

Observation 1215a215-26d3-4149-992f-242c007f4c14 · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.195715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:861e68f617e54875a67eb13aaa0d022abb81009b8bf27cefacc744364b170b04

Observation afd780c0-e976-4d4a-ba33-286272a6b206 · inbound

Reasoning Language Models: A Blueprint cites this paper.

Reasoning Language Models: A Blueprint O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-10T18:36:54.936546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:36:54.936546Z digest=sha256:7e0c36c71a04c96a76fadf3ee48d460aebc35de8547c6f402461b99f9ee51561

Observation 289673d0-3f50-4387-9af6-f85cf2663bfd · inbound

LIMO: Less is More for Reasoning cites this paper.

LIMO: Less is More for Reasoning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:11:37.546358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-17T02:11:36.932541Z digest=sha256:1c31c8ab1a28303057dee5b85708976db47700228858246bba76b5617d67a201

Observation 994c0b76-fd55-4067-8f67-28eb2457c620 · inbound

Step Back to Leap Forward: Self-Backtracking for Boosting Reasoning of Language Models cites this paper.

Step Back to Leap Forward: Self-Backtracking for Boosting Reasoning of Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T00:27:17.655256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:27:17.655256Z digest=sha256:be220e41c439e0b9b95c35a1acccdedf2b4bfa9030a78d45133003458ee3f1b1

Observation 91132f91-c429-4887-8366-b58a00b8def2 · inbound

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools cites this paper.

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T22:03:46.445142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:03:46.445142Z digest=sha256:bbfd3a8f198dcf06946bd7fd3f1c1e9c22b372b23021dd85cc1a0388827a7d40

Observation 654adc1e-7ed5-4d85-901d-6daa106d887e · inbound

Generating Symbolic World Models via Test-time Scaling of Large Language Models cites this paper.

Generating Symbolic World Models via Test-time Scaling of Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T21:45:45.387659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:45:45.387659Z digest=sha256:64cc71a6a8b99b5392b43f738d0dd3e7d377b26b50d47303e4066bd1125027b3

Observation 25dd2f91-ba65-4050-8ee3-fe1df5e68c3c · inbound

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling cites this paper.

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:40:36.017882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:40:36.017882Z digest=sha256:93b7cc3d0ea65235ce425cb832caedfeec10d7e02907d0f2a96962847564c4a5

Observation aa5ce6f6-2465-4fbc-9d8a-af09777caf75 · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.490904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.490904Z digest=sha256:6decaaab79cf6c67ad55182683855bf55302f7b430d0c8b837fbf0159731b559

Observation 5c5fb550-2c52-465f-ab20-67fb06169bfd · inbound

Typhoon T1: An Open Thai Reasoning Model cites this paper.

Typhoon T1: An Open Thai Reasoning Model O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T22:52:54.392447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:52:54.392447Z digest=sha256:74ec0a647229a7f90749bf4201c2f2b8119f5d96a424548ffaf564da332bd192

Observation 88419304-34b7-489b-87a3-160a07a81e7b · inbound

Dynamic Chain-of-Thought: Towards Adaptive Deep Reasoning cites this paper.

Dynamic Chain-of-Thought: Towards Adaptive Deep Reasoning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T20:51:16.718723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:51:16.718723Z digest=sha256:fa79b3d9dfd24946fe663f6bc7b4b9c2858eeebad8f6821541a51b1c8ad062e8

Observation 22993868-7e67-49e7-9c5a-5a96bba0053c · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:36:24.608929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:ac2fa34b56cdb0f77caf45bcceae7acb8ec099fc85882c277be329929e286263

Observation 59a79e5e-ca81-42b3-9bf8-ff8b7c9ab3b0 · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T06:59:03.233406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:b55fca71e041b18583cc79e363a2235e73abf01da59e9ec4caab9d51df72fb0e

Observation 8943fddb-5947-4195-9026-d78d9f388374 · inbound

Generative AI Act II: Test Time Scaling Drives Cognition Engineering cites this paper.

Generative AI Act II: Test Time Scaling Drives Cognition Engineering O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 272

Resolution
unresolved
no resolver link, observed 2026-08-16T12:02:43.679640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:02:43.679640Z digest=sha256:99f8579e81f4964ca0f1bd43c5ba5aa4597a4968d772b200572d6a62deaf96e2

Observation 84bc0180-a944-483a-9f67-666a3dec9ae3 · inbound

ToolRL: Reward is All Tool Learning Needs cites this paper.

ToolRL: Reward is All Tool Learning Needs O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T00:26:48.497341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T00:26:48.291431Z digest=sha256:7177311bb93a96788eb58e0428409ad62333ce1f9795215c19a1f3e085c31502

Observation ee1597cd-1966-4b9e-97a2-1b54d068aae6 · inbound

WebThinker: Empowering Large Reasoning Models with Deep Research Capability cites this paper.

WebThinker: Empowering Large Reasoning Models with Deep Research Capability O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T19:14:25.402991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T19:14:25.283645Z digest=sha256:a08328e3441b66f1e2209320533824e286660e7f754cea897b54860d04619a3a

Observation 71a45054-a0e1-4fe6-9301-eb9f30de0d4e · inbound

Sparks of Tabular Reasoning via Text2SQL Reinforcement Learning cites this paper.

Sparks of Tabular Reasoning via Text2SQL Reinforcement Learning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T10:54:52.605395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:54:52.605395Z digest=sha256:6c66e6157baf2e6ce0f94dc573e98784c32b53b27e84df548b23ca8f97af9491

Observation 4d3d3a5a-e09c-4cc0-b5bb-a76acac8de03 · inbound

Quantitative Analysis of Performance Drop in DeepSeek Model Quantization cites this paper.

Quantitative Analysis of Performance Drop in DeepSeek Model Quantization O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:58:05.765285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:58:05.765285Z digest=sha256:1da7f3ef260483e5df1860b5be5dc61fccb0b973930a5a2f93db3f3b6d7b5aaf

Observation 59a469f5-9a85-4d42-872c-d760af067d3a · inbound

Long-Short Chain-of-Thought Mixture Supervised Fine-Tuning Eliciting Efficient Reasoning in Large Language Models cites this paper.

Long-Short Chain-of-Thought Mixture Supervised Fine-Tuning Eliciting Efficient Reasoning in Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:46.276015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:56:46.276015Z digest=sha256:00a6d79c5b0d44079ca437a965061a9203a8e1a221390c7f627466cf73efc8bf

Observation 76453cb6-ab27-4fb2-a1c0-6d1c1392b019 · inbound

Scalable Chain of Thoughts via Elastic Reasoning cites this paper.

Scalable Chain of Thoughts via Elastic Reasoning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:13:33.312069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:13:33.312069Z digest=sha256:69f9ad13fc729845cead112748352d3899301d5533265d51ceba380743571b60

Observation 1bc18ce3-73d5-4d44-8a5a-84a82dc7d135 · inbound

Parallel Scaling Law for Language Models cites this paper.

Parallel Scaling Law for Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T21:14:45.785477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:14:45.785477Z digest=sha256:a6e375c21bc9f3bab97bcc8bf3fd4941ee7d9e57251b2549c1ada404a6562b5b

Observation e0f2a2b3-80f0-445d-ad20-e698540706d5 · inbound

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models cites this paper.

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:23.212775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:23.212775Z digest=sha256:f7f69f6b6b8a6d7770cb5760bff90b8576ec1d477e212a08a2c18eb6dd3c06a0

Observation cca941a6-7dfc-4252-aaf5-127ec6ba9e79 · inbound

THOR-MoE: Hierarchical Task-Guided and Context-Responsive Routing for Neural Machine Translation cites this paper.

THOR-MoE: Hierarchical Task-Guided and Context-Responsive Routing for Neural Machine Translation O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:40.245671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:40.245671Z digest=sha256:0e31f9a39825d88a703b7cdb680af8daab6b8541424212c105b7c18944f894df

Observation 0adb45c0-3add-4d97-9c99-26536730064d · inbound

Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning cites this paper.

Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:06.440420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:06:06.440420Z digest=sha256:db1d1618417b103b6553c03155213fb33000a5d1d453f9b4caf18ff1e675e407

Observation ddff6edb-6561-415f-a0be-7db1e94857ea · inbound

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL cites this paper.

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:35.110924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:35.110924Z digest=sha256:96b4d6461439fcfcc5029c6ee1f0573e01cf6689630cf369585519cdba310cdc

Observation 8be07af9-7085-4b03-a54c-900c6476ec1a · inbound

ManuSearch: Democratizing Deep Search in Large Language Models with a Transparent and Open Multi-Agent Framework cites this paper.

ManuSearch: Democratizing Deep Search in Large Language Models with a Transparent and Open Multi-Agent Framework O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:15.669194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:39:15.669194Z digest=sha256:cbb36637641b325026757d858a2b80e3b40e60b43e5f2d87eb3f8d27ad3fd251

Observation c47a79b1-30cf-4b43-a5c0-f515f443b4e2 · inbound

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective cites this paper.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:08.930615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:08.930615Z digest=sha256:4096086d21cbe9b31abc5d26987997498b33c26ec2133317b892a658f65f1d46

Observation 226526dd-0686-47ca-ae8f-4bf3a708b26f · inbound

Which Data Attributes Stimulate Math and Code Reasoning? An Investigation via Influence Functions cites this paper.

Which Data Attributes Stimulate Math and Code Reasoning? An Investigation via Influence Functions O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:07:48.828230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:07:48.828230Z digest=sha256:ab334bcb34cde71a3d54f3bc373efe2d7db38870f6c27814dc252fe49d3c3716

Observation 1f7a1448-9caa-4e79-865a-4df6ba198dd8 · inbound

TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment cites this paper.

TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:43:10.814786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:43:10.814786Z digest=sha256:a5039d5ca24fc0861a766017e8aa55be9ca738d2b5084c8003d4e5027b9b9891

Observation 2f5a6936-3892-45ca-925a-64d68330879a · inbound

Discriminative Policy Optimization for Token-Level Reward Models cites this paper.

Discriminative Policy Optimization for Token-Level Reward Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:53:02.180756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:53:02.180756Z digest=sha256:474b4d7ab72f3e9fbae5b330ab6fb84c9ce07f75f9b352500a421704dd2c134b

Observation 7f4e8530-e3d7-4994-a8a2-82b77e6730e9 · inbound

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking cites this paper.

GThinker: Towards General Multimodal Reasoning via Cue-Guided Rethinking O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:13.434001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:13.434001Z digest=sha256:692154800aab484f9f5df89a13f7d679c60d45a92c74c63d317e110ed6308d5d

Observation 2d7ca1f7-6c45-49c7-8011-f26477e5e3b3 · inbound

Unlocking Recursive Thinking of LLMs: Alignment via Refinement cites this paper.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.055033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.055033Z digest=sha256:11ca33ae523dd6200a1e0dc53e0e49cc9eeba6324f604623074e1ec74604e4c7

Observation 3583c1e9-3f48-4c8b-bd63-f60e89578f0d · inbound

Thought Graph Traversal for Test-time Scaling in Chest X-ray VLLMs cites this paper.

Thought Graph Traversal for Test-time Scaling in Chest X-ray VLLMs O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:13.977702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T09:15:46.135084Z digest=sha256:3efaf7749a4deb2a698f25324409c4b46e86379adb8e48af4d606b4481ae56de

Observation 89da2d77-6780-4772-a51a-bf55b651bd75 · inbound

Confucius3-Math: A Lightweight High-Performance Reasoning LLM for Chinese K-12 Mathematics Learning cites this paper.

Confucius3-Math: A Lightweight High-Performance Reasoning LLM for Chinese K-12 Mathematics Learning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:55:51.083275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:55:51.083275Z digest=sha256:d15b654e4f9a09587c24cd989d3ff28c3f374bdf86a22689815861e540117c00

Observation ee4f3b3a-defe-405c-883d-5f473e192fab · inbound

A Survey on Model Extraction Attacks and Defenses for Large Language Models cites this paper.

A Survey on Model Extraction Attacks and Defenses for Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:11.516566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:11.516566Z digest=sha256:c4847c3322723e00ebedc39824fe66981e9b6a9bbc82b45f70a16904cc7a0c2b

Observation fcab31e2-8508-4fe2-9861-2d58435f8910 · inbound

How to Train a Leader: Hierarchical Reasoning in Multi-Agent LLMs cites this paper.

How to Train a Leader: Hierarchical Reasoning in Multi-Agent LLMs O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T18:12:59.282290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:12:59.282290Z digest=sha256:329038b34316ad62925ba051d619de7b503619451b2961ebf602457840091f2c

Observation b8caf060-27c4-413a-b1b9-77cfc1303792 · inbound

Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding cites this paper.

Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:51.208072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:50:51.208072Z digest=sha256:8f3971014583fc8f1b08e4ee00734882b9897169a699fc7c77b84750364be290

Observation 8d9cd287-48e2-4169-96c8-fb4ab79daf80 · inbound

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning cites this paper.

CoLD: Counterfactually-Guided Length Debiasing for Process Reward Models in Mathematical Reasoning O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:25:45.349208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T23:24:43.556606Z digest=sha256:6a0c418083a5bca56df2cdc85a5ebf480243fc6f1b462dea32a90b0dff9320a1

Observation d730a4a7-c698-48cf-9285-c0fd4eb7a813 · inbound

Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory cites this paper.

Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:22:13.256368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:22:13.256368Z digest=sha256:fda90f6c3a988a1f351c4a8e30e9ea1e1ae4a8b4aa9fcd60202e40422abdfb90

Observation 3bdf99f3-2c80-41cc-8a8a-bee263f8a5fa · inbound

Cognitive Duality for Adaptive Web Agents cites this paper.

Cognitive Duality for Adaptive Web Agents O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T23:36:22.055417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:36:22.055417Z digest=sha256:049f7d03bfcecd35641457e7fbcc24c56b2350b396ccb827a4cca7c9ac0a329b

Observation 77f92f80-5b45-4dfb-ab0c-44c68e38d6b9 · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 186

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T19:21:48.787609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:41591233b1037ee676213bdbd87ae78d58e539566c331c87560a91c5e9a98e35

Observation 54025218-799b-439d-b3b8-6be341fb747c · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T10:38:59.862669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:38:59.862669Z digest=sha256:550357e0814749ee2b88cc3491ef9c5ae751a129134d78c092175840d509d746

Observation acc2dabf-d3ea-449f-9358-ffe2e6ef2a66 · inbound

LAMDAS: LLM as an Implicit Classifier for Domain-specific Data Selection cites this paper.

LAMDAS: LLM as an Implicit Classifier for Domain-specific Data Selection O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T23:33:42.000535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:33:42.000535Z digest=sha256:613ae4f747364894f35b59c81b167f3bafe75dc90ea38ea4aeee08d0a2d67851

Observation 362dd49c-e6ae-4bcf-ad29-302e6ba8dfa4 · inbound

AIPO: Learning to Reason from Active Interaction cites this paper.

AIPO: Learning to Reason from Active Interaction O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T08:06:31.700346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T01:17:28.124867Z digest=sha256:3ee7c669ee1d1fa21213da3b8ec8b22a878a1f114164dd496cf798cb6e536237

Observation 6642f3fb-148e-4fc8-a0aa-5eda8559c7a4 · inbound

AIPO: Learning to Reason from Active Interaction cites this paper.

AIPO: Learning to Reason from Active Interaction O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:07:42.310933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T18:07:27.492419Z digest=sha256:93856916317bf0ebfdf2006ec54376d8445e4327b73c6b25c89b35ad4d514e90

Observation e3d227e9-db3c-404e-b0c8-729919741a87 · inbound

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights cites this paper.

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:57:09.849222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T22:16:42.836058Z digest=sha256:f78847563d80e150faee8b760c8d488e46a7eea887e26c26402446882363a97c

Observation 9a7f20ad-195c-471b-9e9b-54f32a386a21 · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 263

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T11:09:46.426975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T08:09:57.542558Z digest=sha256:597c18adf90edfb4c06b556b972c82f0716d8af074ca54dd7d91fbeb0f78dd0b

Observation 606e70da-df3e-4931-8af5-174cabf12691 · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 263

Resolution
unresolved
no resolver link, observed 2026-08-02T10:27:18.616327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:27:18.616327Z digest=sha256:5595b3199a5db299f7c0fa79f24babf3197cca4908ed46315e7d234942fd82e5

Observation 13f33345-a82a-4a91-ad35-95010d2f1d58 · inbound

Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models cites this paper.

Answer-Conditioned Chains of Thought Degrade Verifiable-Reasoning Distillation in Large Language Models O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T01:52:30.041159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:52:30.041159Z digest=sha256:6a5e0716a6654f8312eb09a60373059a78c56ef0a78cd688f704484d938c5b87

Observation a04551ab-739c-444d-9a93-ccf9b4c5ee93 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.505387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.505387Z digest=sha256:df73d1a6fe82a1b3c53181f437059025167ebca9d99f9026bd074ee4eb045327