Pith. sign in

Paper Citation Record · LEDGER

On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2406.12221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.12221 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:47:59.929137Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-09T16:31:18.020108Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 47319f6f-3863-4b98-a068-1f170de20958 · inbound

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering cites this paper.

Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-12T18:28:49.136772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:28:49.136772Z digest=sha256:7c8ff857354739e83ed6537639e8f55eae098b882dac16213c4204713f661972

Observation f37531f3-90eb-41e2-8221-d4cffac59d4e · inbound

DeepRAG: Thinking to Retrieve Step by Step for Large Language Models cites this paper.

DeepRAG: Thinking to Retrieve Step by Step for Large Language Models On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-09T16:31:18.027454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-09T16:31:17.714524Z digest=sha256:c8bc8ec69d04166f5c5b35f65952a3de2155bcced15ef17af4171201e3aa892a

Observation 94e9bbb2-fb5c-44c9-b6d2-e1aa7f4966f6 · inbound

Sailing by the Stars: A Survey on Reward Models and Learning Strategies for Learning from Rewards cites this paper.

Sailing by the Stars: A Survey on Reward Models and Learning Strategies for Learning from Rewards On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T00:47:59.929137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:47:59.929137Z digest=sha256:55b0f3258844f733e2322e01327455c6328522b39f1286df0f24643302a69588