Pith. sign in

Paper Citation Record · LEDGER

Large Language Models for Mathematical Reasoning: Progresses and Challenges

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2402.00157.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.00157 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 58 of 58 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T05:44:40.357587Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T11:39:46.423009Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f7bee619-0adb-4582-aea6-e6643ed9362a · inbound

TS-Reasoner: Domain-Oriented Time Series Inference Agents for Reasoning and Automated Analysis cites this paper.

TS-Reasoner: Domain-Oriented Time Series Inference Agents for Reasoning and Automated Analysis Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:45:47.247483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T19:45:39.130509Z digest=sha256:b45152680f13d42d75bc55e2eba8621c09fdb3618e119d6f38f9bb6aa9c897fa

Observation 13237bbd-396e-47ce-8b20-4238944258aa · inbound

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection cites this paper.

ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:13:24.615341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T20:10:59.264484Z digest=sha256:50dc8d73128555ffeb9141679c03b892e31b0c304ee34086c56f54c167c6b5fd

Observation ea4c6404-004c-4d0b-9fd0-2ec1890770cb · inbound

AL-Bench: A Benchmark for Automatic Logging cites this paper.

AL-Bench: A Benchmark for Automatic Logging Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T05:44:40.357587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T05:44:40.357587Z digest=sha256:2f8f8592b63c836de909fc9a076a957bd2496317bb6e123b14cd7a2eaa2bbf6f

Observation dbe155a3-5eab-4969-8018-e5bfe5a2dc82 · inbound

Schema-Guided Scene-Graph Reasoning based on Multi-Agent Large Language Model System cites this paper.

Schema-Guided Scene-Graph Reasoning based on Multi-Agent Large Language Model System Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T04:45:34.659093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:45:34.659093Z digest=sha256:5ad42ba1c97e8a355dff42fbe8545498bec3bed890d80e9343ec0e966f1c4890

Observation 53b0ae2d-e626-445c-b340-c8e6260c1516 · inbound

Optimizing Temperature for Language Models with Multi-Sample Inference cites this paper.

Optimizing Temperature for Language Models with Multi-Sample Inference Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T20:01:43.661854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:01:43.661854Z digest=sha256:c20dc2bc3433f7b8a2535e02f45466d72d45860bc1280dd59d11536235b2b22f

Observation 3aa80304-020b-4b5b-a6fe-8b0a762a150d · inbound

Enhancing LLM Character-Level Manipulation via Divide and Conquer cites this paper.

Enhancing LLM Character-Level Manipulation via Divide and Conquer Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T10:10:44.347345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T10:10:44.347345Z digest=sha256:8f2ad15da0fcaab5cdd11de00a40f115345b030f289661a80d326caa1c6d31ce

Observation c4a713b8-e40b-4c48-a05b-781ed5aeaa3c · inbound

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization cites this paper.

R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T00:19:20.632524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T00:19:20.462455Z digest=sha256:a398bd2eb1f1700f8fc8324185dd686fa44a29a644c52ca03e97e9fd3051d102

Observation 04501110-030e-4523-9fd0-d0f864796504 · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.238711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:f48e2852628382bd7f9dae8d6b8f1a4c5c4706a748d306cb340e8d28a01fe754

Observation f8c76616-33a6-4103-8028-772e93c789c0 · inbound

Bridging Language Models and Financial Analysis cites this paper.

Bridging Language Models and Financial Analysis Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T01:12:20.566155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T01:08:58.528533Z digest=sha256:502c2e2674c14210d0803948b20db3adfcd962140c6ec46fe80052cdbb5c6df7

Observation 0626d88e-ff78-485b-872c-82391f719e8b · inbound

MIST: A Co-Design Framework for Heterogeneous, Multi-Stage LLM Inference cites this paper.

MIST: A Co-Design Framework for Heterogeneous, Multi-Stage LLM Inference Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T21:17:08.075906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T21:16:31.655330Z digest=sha256:e2bf1ffc03adbc2f193ce9fedc9a79ec29ae42ca611a684b6f9b8f8c6f062757

Observation 0debd0e8-2f95-46c1-922f-d2871a15e295 · inbound

Can reasoning models comprehend mathematical problems in Chinese ancient texts? An empirical study based on data from Suanjing Shishu cites this paper.

Can reasoning models comprehend mathematical problems in Chinese ancient texts? An empirical study based on data from Suanjing Shishu Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:05.646723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:05.646723Z digest=sha256:f7ad0b635f1bcf2b984c53065ee9d03e2a32be276c010d1453b584febcdfd0a9

Observation e9182548-547a-48f2-9993-f7c7ca9a5b69 · inbound

Towards General Continuous Memory for Vision-Language Models cites this paper.

Towards General Continuous Memory for Vision-Language Models Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:42.184846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:42.184846Z digest=sha256:d749c569423e21001791b99cf622ba058f8996ee7d06005b097e3fa454eb5e2c

Observation 11854a30-41ac-4009-8b4a-a13ac9f96c7f · inbound

Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots cites this paper.

Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:32:19.252896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T13:29:34.151546Z digest=sha256:c939ad328afe5d7a787e072af379dfda60f343eb9b29c7d0d2eda3ca38a56fe6

Observation 34ac6422-b4c4-4d98-8f2d-935ee4a39d94 · inbound

CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis cites this paper.

CulFiT: A Fine-grained Cultural-aware LLM Training Paradigm via Multilingual Critique Data Synthesis Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:26.062156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:17:26.062156Z digest=sha256:a38fc098325f2adadb99bd3d3c06d5047cb6cb1cd1d92193c09d047eb295a6b1

Observation 86ee1e6a-fce2-49dd-a44a-16f4bd01b9e5 · inbound

Think Again! The Effect of Test-Time Compute on Preferences, Opinions, and Beliefs of Large Language Models cites this paper.

Think Again! The Effect of Test-Time Compute on Preferences, Opinions, and Beliefs of Large Language Models Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:40.022995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:14:40.022995Z digest=sha256:ce48b101d6aefce67a46b3f48c5ecfbefd2f31e7bd23c113930c388243f09329

Observation 2cc13414-55b7-4a9c-abf4-e85bea0b9a3f · inbound

Probability-Consistent Preference Optimization for Enhanced LLM Reasoning cites this paper.

Probability-Consistent Preference Optimization for Enhanced LLM Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:48:47.051024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:48:47.051024Z digest=sha256:59a701cea4c964086b0fcfba69cd5f01fd4d0d5cc4335b863358abdcbf3c7d11

Observation 6f1d3a81-d2e7-4427-9aac-d007c05be2cc · inbound

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction cites this paper.

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:37:15.930637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T11:34:09.428653Z digest=sha256:9d8fede798d6b755f884854611ad8ef4d793d62ff9d95ea9d4ca783f329f758d

Observation a381e39a-81c4-4b8d-b03a-f5c4c46c26f0 · inbound

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains cites this paper.

Knowledge or Reasoning? A Close Look at How LLMs Think Across Domains Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:33:29.433062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:33:29.433062Z digest=sha256:4fb7f492ba6077f3acfed588678064dccbbdfd1df8318a6de60e1f7be8754be2

Observation 74ae6843-d300-4d16-a475-a45ff3ad223e · inbound

Learning to Insert [PAUSE] Tokens for Better Reasoning cites this paper.

Learning to Insert [PAUSE] Tokens for Better Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:16.657975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:06:16.657975Z digest=sha256:002c086446694df84808b727c9a3b7837da949e594c0c25701c4aef4ecae0c30

Observation 2029e78c-c795-4bb7-9a6e-ae7d46f2f301 · inbound

More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning cites this paper.

More or Less Wrong: A Benchmark for Directional Bias in LLM Comparative Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:59:03.671310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:59:03.671310Z digest=sha256:ada2c9d9af4c68027efb28a85a74ca12b941b3dda42ce26fe339bc71bdf11478

Observation ecd86cd7-3a10-4636-b632-d0e84df0588b · inbound

Structured Pruning for Diverse Best-of-N Reasoning Optimization cites this paper.

Structured Pruning for Diverse Best-of-N Reasoning Optimization Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:56:23.511738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:56:23.511738Z digest=sha256:bf17f1fecf7cb6c8c02f5c4804c79abf72c80b0a5d4e4fd1c79ee5b7eb44364b

Observation 0a8b4406-5832-45e1-8897-10e849757da2 · inbound

Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification cites this paper.

Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:44:59.314638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:44:59.314638Z digest=sha256:4adc1257d243289bc7e410fd4c0d72253891cc8b53d87751ff94376e64541a2f

Observation be2ce9b2-d613-4dbb-a653-1a88b64c37dc · inbound

Foundation Model Empowered Synesthesia of Machines (SoM): AI-native Intelligent Multi-Modal Sensing-Communication Integration cites this paper.

Foundation Model Empowered Synesthesia of Machines (SoM): AI-native Intelligent Multi-Modal Sensing-Communication Integration Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:33:51.500639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:33:51.500639Z digest=sha256:9a5181fc7ea58229356dd4d1d63c34c90af8b0ba831d7a450070257dec583dd6

Observation 24a13888-daca-4c40-a5e3-a013ce5f0b46 · inbound

WIP: Large Language Model-Enhanced Smart Tutor for Undergraduate Circuit Analysis cites this paper.

WIP: Large Language Model-Enhanced Smart Tutor for Undergraduate Circuit Analysis Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:00:32.194052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:00:32.194052Z digest=sha256:fdc55bfbbf53b081be44f23910223d03e326d04f3c325037f6dcd8417968116c

Observation d1d4f725-ebfb-4830-a238-2ba40941af84 · inbound

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions cites this paper.

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:08.038342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:08.038342Z digest=sha256:d5059d40c190489e37f7be3f3d15da2b503db985bc4134992cbec8dbd44c228d

Observation d739b67a-0cd2-485d-a83c-f8f98f421293 · inbound

WGSR-Bench: Wargame-based Game-theoretic Strategic Reasoning Benchmark for Large Language Models cites this paper.

WGSR-Bench: Wargame-based Game-theoretic Strategic Reasoning Benchmark for Large Language Models Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:34:37.342152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:34:37.342152Z digest=sha256:98c151fd54872a13e0c55737bfc71bde87e8cfb0483c44d15a6d2f40529e96e2

Observation a761afb0-eef9-4f53-bd59-c84dae9dabb1 · inbound

Answer-Centric or Reasoning-Driven? Uncovering the Latent Memory Anchor in LLMs cites this paper.

Answer-Centric or Reasoning-Driven? Uncovering the Latent Memory Anchor in LLMs Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:33:52.072107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:33:52.072107Z digest=sha256:ffcaec74d675f02e0c11822770023f7be9cfe80aef079df4dc150a4687caae8f

Observation fa880957-aadb-49d0-983f-bc115dddfae6 · inbound

A Large Language Model-Empowered Agent for Reliable and Robust Structural Analysis cites this paper.

A Large Language Model-Empowered Agent for Reliable and Robust Structural Analysis Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T22:21:09.695912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:21:09.695912Z digest=sha256:868e77638ef5dcee73dbd65d8b85f8bc090feb1f145dd034039d44ed393a042d

Observation 53abae40-7d72-498a-b040-d722d87eaa46 · inbound

Fine-tuning Large Language Model for Automated Algorithm Design cites this paper.

Fine-tuning Large Language Model for Automated Algorithm Design Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:40:46.609814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T23:36:12.949230Z digest=sha256:498f3f199a460efba0e59e71e8a361ff14550c97257693a7e2092e425eb70b8a

Observation 24462251-b2e1-4f83-9d93-4fbf38474a15 · inbound

Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding cites this paper.

Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:50.090392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:50:50.090392Z digest=sha256:a09af7f5a40224ae9a2e76ecaa660daef0e90d401726b8b167151ec2bb72540d

Observation 111c09e7-d199-49e0-b63d-8387169a80df · inbound

League of LLMs: A Benchmark-Free Paradigm for Mutual Evaluation of Large Language Models cites this paper.

League of LLMs: A Benchmark-Free Paradigm for Mutual Evaluation of Large Language Models Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:22:01.377367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T03:17:06.457421Z digest=sha256:5278ed9a406d20de3265699471063fb01b7cde827094543b291da09ba50e9917

Observation 7a9e52fd-f93e-4f87-9c88-d049b7cc2b5e · inbound

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning cites this paper.

GM-PRM: A Generative Multimodal Process Reward Model for Multimodal Mathematical Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:59:51.320079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:59:51.320079Z digest=sha256:21238814776444408060bf0526d5bdf8a96e25fa55fd24feba550d8540ae6517

Observation 62408393-31d2-47ba-943b-972c66d72903 · inbound

Arrows of Math Reasoning Data Synthesis for Large Language Models: Diversity, Complexity and Correctness cites this paper.

Arrows of Math Reasoning Data Synthesis for Large Language Models: Diversity, Complexity and Correctness Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T16:15:28.188254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:15:28.188254Z digest=sha256:967a70137dfd8ac64b7611612dd996c93e067875da63063fd423c4364cf6bc70

Observation 512aafcd-8ce9-4e71-992d-fd284663d72a · inbound

Do AI Models Dream of Faster Code? An Empirical Study on LLM-Proposed Performance Improvements in Real-World Software cites this paper.

Do AI Models Dream of Faster Code? An Empirical Study on LLM-Proposed Performance Improvements in Real-World Software Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:41:00.361974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T06:39:42.391102Z digest=sha256:7b8713bea4805130f308453cffc263955683fe520f10d794ddefac596405bd8c

Observation 4c1eaf67-c7ec-418a-9be1-98e3e788253b · inbound

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle cites this paper.

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:20:58.331885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T06:19:44.360734Z digest=sha256:84c61870d46667422f484b29fd1b967ac31eb4f8d185e23c168ea6c8e8565afa

Observation f5e710a1-c167-47cf-a778-f3f460b7ed4f · inbound

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle cites this paper.

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:50:36.463555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T20:50:06.642976Z digest=sha256:5d0c0fa122c0df732de7db64416d1a7b94306c7b367ebe87352217c00de7a6d9

Observation 5d3077fa-e6e1-4769-a846-722822d633b5 · inbound

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning cites this paper.

ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T17:53:53.227185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:53:53.227185Z digest=sha256:5a11e8263d7371e7bcaa416e71204cadf85820bd8f15ee13f13e665e31b2feef

Observation 422c3237-8120-4ad2-82f6-a6e0c104d21a · inbound

One Tool Is Enough: Reinforcement Learning for Repository-Level LLM Agents cites this paper.

One Tool Is Enough: Reinforcement Learning for Repository-Level LLM Agents Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T14:20:31.208360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:20:31.208360Z digest=sha256:a5a5365e353084bf8d0481ee45230b9ac7975a26ea5d14b301aea0872ffa0573

Observation d299b128-0f29-4165-b051-41baea0c6acf · inbound

From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning cites this paper.

From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T06:52:52.486546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T06:52:52.486546Z digest=sha256:3148fe04d04bf26696ca26abb41861259a3f680885c1a843363af5169135a052

Observation e830e4be-7dd7-4964-a2eb-92129905252c · inbound

Adaptive Information Control for Search-Augmented LLM Reasoning cites this paper.

Adaptive Information Control for Search-Augmented LLM Reasoning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T05:39:26.664534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:39:26.664534Z digest=sha256:374c3299ffc4ab070bdab40478c7d9ed6a49f99f1fa3b403327c795f6fe09858

Observation 334648bc-16d3-4fe1-8a89-be89d1d21707 · inbound

RACC: Representation-Aware Coverage Criteria for LLM Safety Testing cites this paper.

RACC: Representation-Aware Coverage Criteria for LLM Safety Testing Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:17:36.698867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:12:55.296932Z digest=sha256:87a9f7fbf079ecc985f91066a26431ba15edd20a8fb61885358007cbbfbc5d3e

Observation 7f0e494e-5058-4140-a1f9-5b67215432df · inbound

Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals cites this paper.

Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T05:16:27.580989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:16:27.580989Z digest=sha256:e1912477c7dbf548b562afdf0b253400711bec6d77229a81248b6957ca71b0cb

Observation 7e3a12e5-a8f9-42fa-876c-a72a0125441f · inbound

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving cites this paper.

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:16.038853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T06:33:29.860024Z digest=sha256:41cee4455720a2b49c45d6ea2aacfffa0a246de6293f277368dd869f7e9360fc

Observation c949b87c-15f4-489d-a476-0f37a94bcc5f · inbound

Improving Medical VQA through Trajectory-Aware Process Supervision cites this paper.

Improving Medical VQA through Trajectory-Aware Process Supervision Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:06.836267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:18:17.846140Z digest=sha256:97119d0cb136718274c14b8624f4575fef04ba9a879d6db6eab4b7f13d7a6106

Observation d33476eb-029c-42a1-893c-bc56de0cd37b · inbound

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning cites this paper.

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:06:33.167115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:45:06.199636Z digest=sha256:b76e95b503b0dd3c0653e6ab00d465cd3c429fe618e1ad673039f67c01b8cc62

Observation 37686144-eed6-4c86-87a8-55572a54d23d · inbound

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning cites this paper.

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:23:48.380113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:19:49.016156Z digest=sha256:816da88928cb960eb731baed31ebac3a4e36d744f8bbede361be8ab40db9cfda

Observation f3f1dc8c-8d96-48bb-9328-57bff04c810e · inbound

CLORE: Content-Level Optimization for Reasoning Efficiency cites this paper.

CLORE: Content-Level Optimization for Reasoning Efficiency Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:51:08.336753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T05:50:23.111591Z digest=sha256:1963b7a719c5ca6f297ed3578ec6c39340360eb3b39b4d117dd44deb24876dac

Observation c32279f1-64e2-4c06-b8d5-b7cf6048a143 · inbound

Inferring Code Correctness from Specification cites this paper.

Inferring Code Correctness from Specification Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:33:31.522748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T06:33:36.835860Z digest=sha256:0b830e246561ddb452d44dee56d55887daae5938a27fa8e031194126c5d2db69

Observation 3471314e-5241-41d6-a401-14001a03eb74 · inbound

LLM Parameters for Math Across Languages: Shared or Separate? cites this paper.

LLM Parameters for Math Across Languages: Shared or Separate? Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:28:59.014009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T00:30:30.315423Z digest=sha256:aa95c440abb94b0417b22b54138e290516cbb017f3ea1dbe43f42015e84773d0

Observation 87d19f32-075e-41c3-ba18-fe4ca06fb773 · inbound

One Generator, Any Process: LLM-Conditioning for the LHC cites this paper.

One Generator, Any Process: LLM-Conditioning for the LHC Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 244

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T11:39:46.424583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T07:53:57.250401Z digest=sha256:58065c02fe2535d84bca8de2e74ec8c8b4c6d171e83a3e8136b954c54fbcc680

Observation b2e94b4c-eda2-4d4c-9d40-96e85ea54390 · inbound

One Generator, Any Process: LLM-Conditioning for the LHC cites this paper.

One Generator, Any Process: LLM-Conditioning for the LHC Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 248

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:14:36.053955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T10:13:09.503522Z digest=sha256:2514e9660207c5142644bca9976d80c6c73b4f347817ec1ec4e0786f7b6c7540

Observation 41734c81-27c9-49d2-87d2-1667ebe3f3bb · inbound

RareDxR1: Autonomous Medical Reasoning for Rare Disease Diagnosis Beyond Human Annotation cites this paper.

RareDxR1: Autonomous Medical Reasoning for Rare Disease Diagnosis Beyond Human Annotation Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:27:18.591729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T19:21:44.653877Z digest=sha256:9e4afaff1fb71e74cde9b103c2122a101f66208a79a8988a3fa64e98aa2a0741

Observation 1b116a76-8459-4a3c-9886-750be77e14ef · inbound

SCAPE: Accurate and Efficient LLM Training with Extreme Sparse Communication cites this paper.

SCAPE: Accurate and Efficient LLM Training with Extreme Sparse Communication Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:58:46.648992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T17:56:52.510949Z digest=sha256:e28d0a53052d02ff12022e371f14d4013e4894a84eeac982fdb252c0ab8de530

Observation 57ed5482-221d-49bf-8944-8629815e0a52 · inbound

STEC: Evidence Compression for Deep Search in Open-domain Multi-Hop QA cites this paper.

STEC: Evidence Compression for Deep Search in Open-domain Multi-Hop QA Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T09:13:24.763990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T09:13:24.763990Z digest=sha256:7ce2c57d61da04c8fc42aecd1cbd17454a8729aeb4254b78bd25f7ac1e65045a

Observation bf9c8779-e9c0-498a-b366-13b4c4e7c92b · inbound

Feature Generation Using LLMs: An Evolutionary Algorithm Approach cites this paper.

Feature Generation Using LLMs: An Evolutionary Algorithm Approach Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T09:47:29.204988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:47:29.204988Z digest=sha256:04ea232ad0e3d4d13ee4fa9fe3ffc67ec9e89156e92d9989f173fc735ede79a2

Observation 3412a780-27cc-4607-a29d-79e7642dda2a · inbound

Representation Robustness Under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving cites this paper.

Representation Robustness Under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T08:02:56.456065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:02:56.456065Z digest=sha256:21d1d9aaf3153e9ddc102358956e22b5113c890c388a63cfd8afacd6789b55ea

Observation 97a5f922-c734-4b23-81aa-f36ddadd01c2 · inbound

Assessing the Benefits of Combining Advanced Deep Learning Techniques for Post-Disaster Building Damage Assessment from UAV Imagery cites this paper.

Assessing the Benefits of Combining Advanced Deep Learning Techniques for Post-Disaster Building Damage Assessment from UAV Imagery Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T18:33:54.681432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:33:54.681432Z digest=sha256:c8dc0f766a17a798ec14e3ce46bb9211e596560c4da98c215241decb111e86d0

Observation 48d54ee0-29f6-4edc-8a3a-0ff5dfb28b6c · inbound

Superloop Equations and Minimal Surfaces I: Confining minimal surface in $4D, N=1$ SYM cites this paper.

Superloop Equations and Minimal Surfaces I: Confining minimal surface in $4D, N=1$ SYM Large Language Models for Mathematical Reasoning: Progresses and Challenges

Reference 277

Resolution
unresolved
no resolver link, observed 2026-08-04T10:28:20.864753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:28:20.864753Z digest=sha256:9d179525e6e9a1f7e54c75e1cce6b3379e19c31425edc3f296570ff95b6c7e36