Pith. sign in

Paper Citation Record · LEDGER

LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2404.05221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05221 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:43.032974Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:37:36.761948Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ea797387-5957-4062-a935-6220882c75de · inbound

Training Large Language Models to Reason in a Continuous Latent Space cites this paper.

Training Large Language Models to Reason in a Continuous Latent Space LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:29:05.695899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-11T10:29:05.384381Z digest=sha256:0831e855d03d7ec82fbfd4a4a9f232d2fc3751f9866cea6e6f07d193d6806078

Observation 5a350cd1-9f03-44ba-9458-0bbdcf69b8ce · inbound

Psychometric-Based Evaluation for Theorem Proving with Large Language Models cites this paper.

Psychometric-Based Evaluation for Theorem Proving with Large Language Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T17:36:09.628897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:36:09.628897Z digest=sha256:55b4493f7031bc23b886e57ac9b297eab9583e1a6d14b851c82093fa0d0dbd3c

Observation 9e76c176-aa0a-4cfc-bfd0-464c0244565c · inbound

Policy Guided Tree Search for Enhanced LLM Reasoning cites this paper.

Policy Guided Tree Search for Enhanced LLM Reasoning LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T11:20:31.771367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:20:31.771367Z digest=sha256:849f5b7e921f32ebd3adb696db43b264dc1d83e4bc85acc2dece052928f2e0b0

Observation d72cb52d-a215-4e79-b62f-e701f4e04ec2 · inbound

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs cites this paper.

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:03.338525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:01:03.338525Z digest=sha256:70965704a30761d5f3df2c9dffd6a076a7a9fc1ead67bd4ae0a0506aa9045c61

Observation 06dd67a5-d55f-47c5-b36b-ee77a3530bc6 · inbound

Generative AI Act II: Test Time Scaling Drives Cognition Engineering cites this paper.

Generative AI Act II: Test Time Scaling Drives Cognition Engineering LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-16T12:02:43.032974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:02:43.032974Z digest=sha256:a151665e1e48b1557009ef967cbe36215c92af3879673c0f22e47ef97d31f362

Observation 83f28c52-37e4-42fd-abdb-b02d042ec52e · inbound

Evaluating Intermediate Reasoning of Code-Assisted Large Language Models for Mathematics cites this paper.

Evaluating Intermediate Reasoning of Code-Assisted Large Language Models for Mathematics LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:38:27.817668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:38:27.817668Z digest=sha256:50729747a0da26c91cbb1a98891aff083b434fb4d910717b0bf1d2e07d8a2ab3

Observation a2ac3aeb-bb24-43f5-b8eb-1e400fb35b00 · inbound

MINERVA: Evaluating Complex Video Reasoning cites this paper.

MINERVA: Evaluating Complex Video Reasoning LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.072960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.072960Z digest=sha256:20dda676b91c5dd8473a33b62a7203d4fdb342ca3471bc031b35ce40665dad60

Observation 9569c7ac-e4c5-4d51-897e-608b72ad4f3a · inbound

DSADF: Thinking Fast and Slow for Decision Making cites this paper.

DSADF: Thinking Fast and Slow for Decision Making LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T22:05:50.975690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:05:50.975690Z digest=sha256:70e65ac6031a7b7abf7731f313f38b94a608b67c2a1a1f59beb34c242c481f07

Observation 55a30775-b88e-4659-869c-6cd87d0f5543 · inbound

CoIn: Counting the Invisible Reasoning Tokens in Commercial Opaque LLM APIs cites this paper.

CoIn: Counting the Invisible Reasoning Tokens in Commercial Opaque LLM APIs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:15:53.715997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:15:53.715997Z digest=sha256:b5603eadfdd533e213839cb7fb4f9111b0cc455830a9b948afd8f615a69988b0

Observation 133cf5fd-32ee-407d-9e8a-9c7fd5dcb72c · inbound

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? cites this paper.

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:21.390374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:21.390374Z digest=sha256:3e79bd12af70acecab9081166d1711e0a4f2263e731dc0cebf61685d80cc2785

Observation e3725d2b-e79d-425a-8ce4-15a9d4f1ad1a · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.814532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.814532Z digest=sha256:41b59233d222f02d2ae68762e3313c24ce929805c3a71889a21aef6e78ba9ef5

Observation ccb135e9-5289-4e47-a40a-7ac2e6897340 · inbound

General Agentic Planning Through Simulative Reasoning with World Models cites this paper.

General Agentic Planning Through Simulative Reasoning with World Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T12:41:33.604022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-22T12:37:10.956233Z digest=sha256:95f0b9e72ba316d7e03d12fa702bfb116c420837cae6e1bd2df89fc47c0a02f5

Observation be4baa25-070f-4361-9342-cea5481a0d48 · inbound

CodeGrad: Integrating Multi-Step Verification with Gradient-Based LLM Refinement cites this paper.

CodeGrad: Integrating Multi-Step Verification with Gradient-Based LLM Refinement LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T21:12:43.061768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:12:43.061768Z digest=sha256:83c806a36155a3fdc07f93c157aa2d156a2268357b64af565e48b009b0cfb28e

Observation 3fdeaaf3-9c18-4fbf-8118-b7216bb93aee · inbound

Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction cites this paper.

Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:26:47.249697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-18T19:26:35.001429Z digest=sha256:71845119712b85afcf97c82e08a87b1a8fdc95d46018e580d845c84e6a53f4e8

Observation 6f9d87b1-3ff6-4748-b7ec-2b9c192057a0 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:14:25.954129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:bac990d03027ba2f6dfc886460445269ceda2972d27e9a74e36f98c70e87be82

Observation 4474bfe1-7b11-423a-8ee5-6cbcd244ee39 · inbound

Vision-aligned Latent Reasoning for Multi-modal Large Language Model cites this paper.

Vision-aligned Latent Reasoning for Multi-modal Large Language Model LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:50:44.249261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T07:48:23.272002Z digest=sha256:7d5c9d5fe9a75f8aa19422c446c042106c7646bc5f2ad70d0b52ae8793bb5e01

Observation e1b30e5a-163b-47d6-9f63-8cadaa259541 · inbound

MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation cites this paper.

MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.125418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-28T22:24:25.409345Z digest=sha256:5525b56182d8136a5c8286336b3041da9cb99f4be01baf5edcb92e34206fa32f

Observation 5e442ee5-5775-483e-96f9-05734fda89a4 · inbound

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction cites this paper.

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:37:36.764461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T13:48:39.832057Z digest=sha256:e4dd2e133439119a19d7441ef032ad510908f84c5e500cb264b7a7afa78ceca8

Observation 6e397313-2a7f-40a2-ba99-ecb1bed2fad3 · inbound

SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game cites this paper.

SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:35:58.044642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-29T01:56:22.856352Z digest=sha256:2a7aa20f46203becdf1bab1aae30a866f3b39947280c44cfe21f657db91f7cb9

Observation f47ff9a2-068c-4e64-be0e-9c7b000b4700 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:affbad214bc83c1c0afecf6ddf472c1d6977d1b973e366c04d57b9297bc8cc4a