Pith. sign in

Paper Citation Record · LEDGER

Alignment-Aware Decoding

As of 17 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 1 inbound Pith citation observation for arXiv:2509.26169.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.26169 v2

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:40:16.013589Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-25T05:01:31.560963Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T05:05:23.205865Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d316683d-ada4-4c43-8290-744f985fccad · outbound

This paper cites Concrete Problems in AI Safety.

Alignment-Aware Decoding Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.147866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.147866Z digest=sha256:50f591f911215f562fff8fccecd9ef63c27a35be93ebd0e5427996657d01e854

Observation df4c1e42-9008-4fed-8370-14f77c46d3c3 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

Alignment-Aware Decoding KTO: Model Alignment as Prospect Theoretic Optimization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.259926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.259926Z digest=sha256:4bd7169b539f92ba57db5249aa6b8bbfe1f9cd72f665d8d946a3032f6f6a521d

Observation 0d904953-7921-453f-8eab-5a346f6452e4 · outbound

This paper cites Google Nest Learning Thermostat.

Alignment-Aware Decoding Google Nest Learning Thermostat

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:16.013589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:16.013589Z digest=sha256:79ee6e056762d2723464bbafcdaac3cb18792b46e8795a87bbf889b108427907

Observation 7ef8deae-82ea-4b26-a27d-8172c8ca98cb · outbound

This paper cites Proximal Policy Optimization Algorithms.

Alignment-Aware Decoding Proximal Policy Optimization Algorithms

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.564742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.564742Z digest=sha256:9de6701c1a69cc6dc3cf051d2871f533dbb15e18f02e2b8c46db9ee22fb41de3

Observation 64d56e6a-dd81-4671-a9c0-effce24de903 · outbound

This paper cites Se- lective preference optimization via token-level reward function estimation.arXiv preprint arXiv:2408.13518,.

Alignment-Aware Decoding Se- lective preference optimization via token-level reward function estimation.arXiv preprint arXiv:2408.13518,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.637587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.637587Z digest=sha256:95ad9d841bf6e44f77218703106aab3b15d0f609e51c446f157b2e5de112c438

Observation 6d45353d-31f5-45da-842c-4dfe6de29502 · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

Alignment-Aware Decoding SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.787091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.787091Z digest=sha256:7232b60349a1fad966e9c62ea01ac012a2a59d6be60a8b7d70cfa41f468d498a

Observation e77889db-73b4-4326-9fee-c156a82bd97c · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Alignment-Aware Decoding Fine-Tuning Language Models from Human Preferences

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.868294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.868294Z digest=sha256:f114aee2e2c746e5ddc7fc5ffe872750de01c5b69e7f2181f0978b7967db6373

Observation ba7d2237-8d97-4449-875b-680d12a10d05 · outbound

This paper cites Table 3: Accuracy of the reward models trained on the different preference datasets.

Alignment-Aware Decoding Table 3: Accuracy of the reward models trained on the different preference datasets

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.959617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.959617Z digest=sha256:840944bdaaca0411b90f5bc518d31f19b7d9367cd5187ebd85235fe816076d05

Observation 13fbb65e-4c45-47fc-a849-db152d582a64 · outbound

This paper cites Contrastive decoding: Open-ended text generation as optimization.

Alignment-Aware Decoding Contrastive decoding: Open-ended text generation as optimization

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.473298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.473298Z digest=sha256:ec063174c99ad862443a9f854704f5c139c9a8e23f898373bf6744883c7300ac

Observation 971ecdf0-564c-4bb9-bb34-ffc2c90c8310 · outbound

This paper cites Orpo: Monolithic preference optimization without refer- ence model.

Alignment-Aware Decoding Orpo: Monolithic preference optimization without refer- ence model

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.309502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.309502Z digest=sha256:3572265f30a9e383162fa008d2f21eac2f43283855c7d0bf2c40303073a88762

Observation a51b33d3-29fc-441a-a105-3c8ae1b984a1 · outbound

This paper cites Deal: Decoding-time alignment for large lan- guage models.arXiv preprint arXiv:2402.06147,.

Alignment-Aware Decoding Deal: Decoding-time alignment for large lan- guage models.arXiv preprint arXiv:2402.06147,

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.431897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.431897Z digest=sha256:ea0ae260dfa728af85331c507700a529dc4e78c1b9629910014cc43d05167be8

Observation 691482fb-79d5-4aa8-95bc-5eb99a055d16 · outbound

This paper cites Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37,.

Alignment-Aware Decoding Iterative reasoning preference optimization.Advances in Neural Information Processing Systems, 37,

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.533702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.533702Z digest=sha256:1d3ef4ffa698d71ecd1f289a26208a3fc6306b8d684a3cd2c987f8507191ac1f

Observation 1b70328b-7990-4f70-acbf-6973a9395ae5 · outbound

This paper cites Reward-augmented decoding: Efficient controlled text generation with a unidirectional reward model.

Alignment-Aware Decoding Reward-augmented decoding: Efficient controlled text generation with a unidirectional reward model

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.206137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.206137Z digest=sha256:dfa676c9443364ab511c1288bf910ca6271ddb18c8f1c2f68d32ff741158f39b

Observation d0025a5e-367b-4786-aaf3-5eb8a14f0b2b · outbound

This paper cites Inference-time alignment in continuous space.arXiv preprint arXiv:2505.20081,.

Alignment-Aware Decoding Inference-time alignment in continuous space.arXiv preprint arXiv:2505.20081,

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.719591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.719591Z digest=sha256:5ab6078db30377d5e9d62f4b52cd00ddf030fa7bc0fa06babf228fed771d3a89

Observation fdfada74-1766-4c5c-833e-d55a4ae98e87 · outbound

This paper cites Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen.

Alignment-Aware Decoding Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T13:40:15.355726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:40:15.355726Z digest=sha256:c8fb802d7edbefc5c731452457a054bb5edb929570c0e5cd8e36498f95b3cd62

Pith citing papers

Observation e64bf062-a3e1-4465-9d55-55a4176a46fb · inbound

Convex Optimization for Alignment and Preference Learning on a Single GPU cites this paper.

Convex Optimization for Alignment and Preference Learning on a Single GPU Alignment-Aware Decoding

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-03T13:05:38.669402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-25T05:01:31.560963Z digest=sha256:c6bdbec8034e929424eab806d6eecdf9a4e20e78bdfc4c89b7542e6c0f15bf31