Pith. sign in

Paper Citation Record · LEDGER

LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2404.05221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05221 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:43.032974Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:37:36.761948Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ea797387-5957-4062-a935-6220882c75de · inbound

Training Large Language Models to Reason in a Continuous Latent Space cites this paper.

Training Large Language Models to Reason in a Continuous Latent Space LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:29:05.695899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T10:29:05.384381Z digest=sha256:2b17fa73cacc0c36c3d4c3be0175dc2ecf7ca7c3f3feeca1e6f9120b5ec42d4e

Observation 5a350cd1-9f03-44ba-9458-0bbdcf69b8ce · inbound

Psychometric-Based Evaluation for Theorem Proving with Large Language Models cites this paper.

Psychometric-Based Evaluation for Theorem Proving with Large Language Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T17:36:09.628897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:36:09.628897Z digest=sha256:55b4493f7031bc23b886e57ac9b297eab9583e1a6d14b851c82093fa0d0dbd3c

Observation 9e76c176-aa0a-4cfc-bfd0-464c0244565c · inbound

Policy Guided Tree Search for Enhanced LLM Reasoning cites this paper.

Policy Guided Tree Search for Enhanced LLM Reasoning LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T11:20:31.771367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:20:31.771367Z digest=sha256:849f5b7e921f32ebd3adb696db43b264dc1d83e4bc85acc2dece052928f2e0b0

Observation d72cb52d-a215-4e79-b62f-e701f4e04ec2 · inbound

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs cites this paper.

One Example Shown, Many Concepts Known! Counterexample-Driven Conceptual Reasoning in Mathematical LLMs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:03.338525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T11:01:03.338525Z digest=sha256:70965704a30761d5f3df2c9dffd6a076a7a9fc1ead67bd4ae0a0506aa9045c61

Observation 06dd67a5-d55f-47c5-b36b-ee77a3530bc6 · inbound

Generative AI Act II: Test Time Scaling Drives Cognition Engineering cites this paper.

Generative AI Act II: Test Time Scaling Drives Cognition Engineering LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-16T12:02:43.032974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:02:43.032974Z digest=sha256:a1a9bcb637e6441bdffd7defa169f06f2145140622e0c1bee3cc792f4b571ab8

Observation 83f28c52-37e4-42fd-abdb-b02d042ec52e · inbound

Evaluating Intermediate Reasoning of Code-Assisted Large Language Models for Mathematics cites this paper.

Evaluating Intermediate Reasoning of Code-Assisted Large Language Models for Mathematics LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:38:27.817668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:38:27.817668Z digest=sha256:bd0a19c0aa144a64b951101e02fc7b420996898e290e96372ad9070605179b13

Observation a2ac3aeb-bb24-43f5-b8eb-1e400fb35b00 · inbound

MINERVA: Evaluating Complex Video Reasoning cites this paper.

MINERVA: Evaluating Complex Video Reasoning LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.072960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.072960Z digest=sha256:20dda676b91c5dd8473a33b62a7203d4fdb342ca3471bc031b35ce40665dad60

Observation 9569c7ac-e4c5-4d51-897e-608b72ad4f3a · inbound

DSADF: Thinking Fast and Slow for Decision Making cites this paper.

DSADF: Thinking Fast and Slow for Decision Making LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T22:05:50.975690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:05:50.975690Z digest=sha256:70e65ac6031a7b7abf7731f313f38b94a608b67c2a1a1f59beb34c242c481f07

Observation 55a30775-b88e-4659-869c-6cd87d0f5543 · inbound

CoIn: Counting the Invisible Reasoning Tokens in Commercial Opaque LLM APIs cites this paper.

CoIn: Counting the Invisible Reasoning Tokens in Commercial Opaque LLM APIs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:15:53.715997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:15:53.715997Z digest=sha256:5150802b11257dfc9d42c8ad3bc08719be58815f2c4cd690eefbe60783f25cdf

Observation 133cf5fd-32ee-407d-9e8a-9c7fd5dcb72c · inbound

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? cites this paper.

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:21.390374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:21.390374Z digest=sha256:3e79bd12af70acecab9081166d1711e0a4f2263e731dc0cebf61685d80cc2785

Observation e3725d2b-e79d-425a-8ce4-15a9d4f1ad1a · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.814532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.814532Z digest=sha256:41b59233d222f02d2ae68762e3313c24ce929805c3a71889a21aef6e78ba9ef5

Observation ccb135e9-5289-4e47-a40a-7ac2e6897340 · inbound

General Agentic Planning Through Simulative Reasoning with World Models cites this paper.

General Agentic Planning Through Simulative Reasoning with World Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T12:41:33.604022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T12:37:10.956233Z digest=sha256:df70b95fd6c1f14f6bed4fda8119f7fb1596a97169468ecd2ea8f0366726b835

Observation be4baa25-070f-4361-9342-cea5481a0d48 · inbound

CodeGrad: Integrating Multi-Step Verification with Gradient-Based LLM Refinement cites this paper.

CodeGrad: Integrating Multi-Step Verification with Gradient-Based LLM Refinement LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T21:12:43.061768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:12:43.061768Z digest=sha256:ea267836ff94b57ca9ce6856485cb01d7891fc59964808d3658e2fe3cafcf93e

Observation 3fdeaaf3-9c18-4fbf-8118-b7216bb93aee · inbound

Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction cites this paper.

Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:26:47.249697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T19:26:35.001429Z digest=sha256:6ead04bf45448a0992c3545832ef1f1d45a3c66be78ec1af9c4ace490c44b559

Observation 6f9d87b1-3ff6-4748-b7ec-2b9c192057a0 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:14:25.954129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:7019c8dfaebaff5eea1527c8e508bca356541776929dea8274634206f327c330

Observation 4474bfe1-7b11-423a-8ee5-6cbcd244ee39 · inbound

Vision-aligned Latent Reasoning for Multi-modal Large Language Model cites this paper.

Vision-aligned Latent Reasoning for Multi-modal Large Language Model LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:50:44.249261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T07:48:23.272002Z digest=sha256:7c409fb3f817f20068c1dea66dd7104c6e6b60477c566e218df3c4cb2dc65120

Observation e1b30e5a-163b-47d6-9f63-8cadaa259541 · inbound

MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation cites this paper.

MoG: Mixture of Experts for Graph-based Retrieval-Augmented Generation LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.125418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T22:24:25.409345Z digest=sha256:31af0322eb91630790ec9ec4534a5b925d933d6b9ed8370deebe46bc5c732993

Observation 5e442ee5-5775-483e-96f9-05734fda89a4 · inbound

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction cites this paper.

Supervised Fine-tuning with Synthetic Rationale Data Hurts Real-World Disease Prediction LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:37:36.764461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T13:48:39.832057Z digest=sha256:7a8533a44b80186d55d5e51cf46d7353b91f7881409da9e9fcf30d488c4852d1

Observation 6e397313-2a7f-40a2-ba99-ecb1bed2fad3 · inbound

SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game cites this paper.

SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:35:58.044642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-29T01:56:22.856352Z digest=sha256:0cd51889893a85003ebd8b501c2ab1b56736a980a196b11a259ec937fdc659a4

Observation f47ff9a2-068c-4e64-be0e-9c7b000b4700 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs LLM Reasoners: New Evaluation, Library, and Analysis of Step-by-Step Reasoning with Large Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:affbad214bc83c1c0afecf6ddf472c1d6977d1b973e366c04d57b9297bc8cc4a