Pith. sign in

Paper Citation Record · LEDGER

Training Language Models to Generate Text with Citations via Fine-grained Rewards

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2402.04315.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.04315 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:57:07.616813Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T21:08:25.887635Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a27c65ae-2341-40df-bf92-4389f7cc90c1 · inbound

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey cites this paper.

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:08:25.890876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T21:08:11.787013Z digest=sha256:a090a111f72567bb45d856082843ebfc626f4a760e2fd529d925fa7853c38a3e

Observation e327d508-df15-4fcf-b359-77f8e7798027 · inbound

SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models cites this paper.

SelfCite: Self-Supervised Alignment for Context Attribution in Large Language Models Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T20:57:07.616813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:57:07.616813Z digest=sha256:0a271dbb47b55491d91f2bd3699da1fa2e20817a0c7a060f393b5587c5fcf3d2

Observation 45617367-ef31-4fb2-9d91-6a7148d34bca · inbound

Lessons from Training Grounded LLMs with Verifiable Rewards cites this paper.

Lessons from Training Grounded LLMs with Verifiable Rewards Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:48.358493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:48.358493Z digest=sha256:675298c44fd63b52d1ce95a2e6d9e3dc64a93356479ac682de48d75d5af4b23c

Observation 1d3dd4eb-ea7d-4ae0-af3f-ede33479728b · inbound

Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models cites this paper.

Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:42:09.116389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T07:39:55.633854Z digest=sha256:d06ef2aa59d336d7f5df95f2620fb9649135ce7e3a22a949212ffba6dc295bba

Observation a6831e99-0c4d-4a1a-a604-6fc20f69371b · inbound

Context Attribution with Multi-Armed Bandit Optimization cites this paper.

Context Attribution with Multi-Armed Bandit Optimization Training Language Models to Generate Text with Citations via Fine-grained Rewards

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:22:09.100951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T07:21:20.893732Z digest=sha256:2c6e1a40fa4cea996d231b23c66908fc471aecf2a8941e7ca0564625f5cf9ed8