Pith. sign in

Paper Citation Record · LEDGER

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2411.16579.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16579 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:00:38.405708Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.141555Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5b592302-f65d-4675-a45c-b0b19ad6b29d · inbound

TrimR: Verifier-based Training-Free Thinking Compression for Efficient Test-Time Scaling cites this paper.

TrimR: Verifier-based Training-Free Thinking Compression for Efficient Test-Time Scaling Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:00:38.405708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:00:38.405708Z digest=sha256:176f11c709e5720e50f340324d107c46b0314cf3697402f49b0a14c12b0b096f

Observation e23b3e40-b19e-44cb-9332-01a0262f1503 · inbound

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning cites this paper.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.417118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.417118Z digest=sha256:1708790b2e07b6ffef3fd44d42e1a0b467dd66b70e0132b19934e1a26e78cc3d

Observation 1f10f4ed-9f2c-40c9-bd52-d9d96a11b99d · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 289

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:10.512934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:10.512934Z digest=sha256:e5dc3387e7359a67360356ead4345c17e3a7027f1fff512469b7fa66670a47cd

Observation 1e88dd56-b72d-4bbb-a544-7ee3bc701ca1 · inbound

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving cites this paper.

Step-Wise Formal Verification for LLM-Based Mathematical Problem Solving Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:48:59.112150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:48:59.112150Z digest=sha256:53c9640c0b5b5c90e55c64468ba9a48ae7a606948d480a96e8797d5d309071b9

Observation 9615e3cf-cb0d-45df-8835-0d74c2c1830c · inbound

Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning cites this paper.

Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:30:47.251015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:30:47.251015Z digest=sha256:70e35d4d6f2d21e712a831aaeb8fa52679e916691539f92e156136f6cb3835e4

Observation 83c9f6db-1dba-49fa-94ed-9c3d312980ed · inbound

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems cites this paper.

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T14:58:28.118937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:58:28.118937Z digest=sha256:4feae09298805da719d9f52d966431c79fd30d25fcd6f8ab3297196fdd252789

Observation 7aacda37-b046-471e-be59-73e724a8bec2 · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 171

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:23:14.944730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:81852322e63f7e4cf440d5cd6188d03ebdd95968e8b09ddbc7c42837e23cb5ff

Observation 3396766a-ad2a-4185-b315-f00feb8a3e76 · inbound

Learning from Natural Language Feedback for Personalized Question Answering cites this paper.

Learning from Natural Language Feedback for Personalized Question Answering Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:56:53.045875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:55:16.228037Z digest=sha256:7764ea339eee4f4ef61e126ea34f3b6e65612a62ec31a1c6395173a9dbd56d01

Observation 6eafc153-32dd-404d-9036-6b9ae1198fbe · inbound

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation cites this paper.

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 1992

Resolution
unresolved
no resolver link, observed 2026-08-04T11:11:12.780313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:11:12.780313Z digest=sha256:8c6383fd8e32f02e3e2f3226f63c0c18ad4e948bb7134df41d69ac61f3686672

Observation f5360fa2-2d32-4f45-aa09-c34df55eb6aa · inbound

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning cites this paper.

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:03:04.330035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T16:01:48.789986Z digest=sha256:c92e9c1977a40ca7cb31c48e78cb5982cf12662854d10b5f8d2e51e2c9716fd7

Observation 6207f2ce-2e63-428e-8042-23e0caf6e0ae · inbound

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning cites this paper.

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:49:00.673885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T20:47:16.236629Z digest=sha256:0478488ece9feb455e8df7a5551e330b78d8cd739131127e3effa38474860a65

Observation 39681fe5-9020-4b3e-8b23-27317878f03d · inbound

A History-Aware Visually Grounded Critic for Computer Use Agents cites this paper.

A History-Aware Visually Grounded Critic for Computer Use Agents Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.143173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T13:20:32.432002Z digest=sha256:c97ffd8aa29d9a5d2f405f359d4fc2d5f5af2d31387543f26c9a18824352f6ac