Pith. sign in

Paper Citation Record · LEDGER

Decoding-time Realignment of Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2402.02992.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.02992 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:10:45.868875Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T09:42:04.132913Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation aac42a07-d3a1-4055-a9ed-806690af5668 · inbound

Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents cites this paper.

Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents Decoding-time Realignment of Language Models

Reference 220

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:42:04.137474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-20T09:41:59.979595Z digest=sha256:69a1b1e4813911cc1f975f867e58f0d4394f8768359080e198f7454ee964272c

Observation 0ae139a1-50fd-426e-833d-3e0c37fbb40f · inbound

PILAF: Optimal Human Preference Sampling for Reward Modeling cites this paper.

PILAF: Optimal Human Preference Sampling for Reward Modeling Decoding-time Realignment of Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T23:03:53.148096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T23:03:53.148096Z digest=sha256:3d6a6ec4667cdf0dd9e157fa4dca5506321c53ae2ed77e6a0dd9350e56235d40

Observation 89380ab3-3786-4332-9388-c281404a96d8 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models Decoding-time Realignment of Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:44.356519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:44.356519Z digest=sha256:e268c141c006d42396a98c1b90c46340a14a27a08605af201c9e21c40f09349d

Observation 9886b70b-d619-4207-a3f0-724816eaea85 · inbound

T-POP: Test-Time Personalization with Online Preference Feedback cites this paper.

T-POP: Test-Time Personalization with Online Preference Feedback Decoding-time Realignment of Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:11.958888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:11.958888Z digest=sha256:74331f1afd8262b21ba06d836c105b8c950c72d61796e71b871e6b87bf9ab69c

Observation 136895b3-952e-41b6-8241-e64f54cc20fb · inbound

Representation-Based Exploration for Language Models: From Test-Time to Post-Training cites this paper.

Representation-Based Exploration for Language Models: From Test-Time to Post-Training Decoding-time Realignment of Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T10:09:03.075477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:09:03.075477Z digest=sha256:a1c54302e9d529e1c417d46599be02782e4348975bde51e5f0e9778a9230cf2c

Observation eb9ab07c-46c3-40e2-a422-4d43991d2d82 · inbound

Step-level Denoising-time Diffusion Alignment with Multiple Objectives cites this paper.

Step-level Denoising-time Diffusion Alignment with Multiple Objectives Decoding-time Realignment of Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:15:25.891004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T13:15:13.005136Z digest=sha256:e9807bf6d7b8859ad866b4ced73f513047cb912fe315c6ef30d057d6429ac03d

Observation b1c3566b-70a9-4552-96fa-d9bf4a3bce68 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Decoding-time Realignment of Language Models

Reference 197

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.868875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.868875Z digest=sha256:e24167f130830964d07ff97e11f16f96b78a492919184222eeb2c4aff15c22b0