Pith. sign in

Paper Citation Record · LEDGER

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

As of 21 August 2026, this Paper Citation Record lists 7 of 7 outbound references and 0 inbound Pith citation observations for arXiv:2605.28010.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.28010 v1

Coverage vector

measured 7 of 7 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:10:43.791548Z

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

7 of 7 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 303557f4-bbf9-4ae9-822e-0213d69acacc · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Training Verifiers to Solve Math Word Problems

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.642494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:9e828915073fd1afa0716d441cb7dc349433789e7c075a9d71d41ee3d45fa0b2

Observation f1a743ca-44a0-4527-8789-57adf9d71484 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.644921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:22291c4e5ca00a854c9c12bf13132ba06f54f458617a8e54c40005a075460f11

Observation 820926d8-3a0d-43c6-aef3-1709b9d356fa · outbound

This paper cites OpenAI o1 System Card.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback OpenAI o1 System Card

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.635435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:90cb8a2c98e2b1e522970dea4bad9e7fffafda39f7b22420a73c4f059db1a7e1

Observation b5645702-7a32-41bb-8a32-7793d0360542 · outbound

This paper cites Spice: Self-play in corpus environments improves reasoning.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Spice: Self-play in corpus environments improves reasoning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.640194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:fc8278809cf4f3d67403051ea1234736cbfa160bd633c9cab248d28fdd53099d

Observation 7af0f012-17b3-4b2e-af7e-23559914b313 · outbound

This paper cites InAdvances in Neural Information Processing Systems (NeurIPS), volume 37.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback InAdvances in Neural Information Processing Systems (NeurIPS), volume 37

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T12:10:43.791548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:0c2e07c303d2b282a82766eae3e2dd86d7b852d8cf9dec23aa28b445334ba542

Observation f94c1dde-40e0-4185-8ee6-f222e5c86b01 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Proximal Policy Optimization Algorithms

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.647159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:ab02b9aca73ed583b1109006cc1a99a2122fd1afbda8194b973c8e0a3d3eeaeb

Observation 63f226a2-7a55-4e16-b448-2faf98a44964 · outbound

This paper cites Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.637785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:c34badfe1972cf0e0c3a22c1be76ed668f2ae700891e44435371a590b8c44328

Pith citing papers

No inbound Pith citation observations are available.