Pith. sign in

Paper Citation Record · LEDGER

Impact of Preference Noise on the Alignment Performance of Generative Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2404.09824.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.09824 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:27:54.719067Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:17:08.587908Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 20f45778-8824-4e59-b186-39645a47c3e7 · inbound

How Humans Help LLMs: Assessing and Incentivizing Human Preference Annotators cites this paper.

How Humans Help LLMs: Assessing and Incentivizing Human Preference Annotators Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:57:29.618598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T03:56:18.703995Z digest=sha256:595fea8475f6c9c2d784a769f0e2d810b0a29b90c4df9986dd8d7c2fe63656a0

Observation da9e0280-38fe-45bd-b58c-f5fbf7c727d1 · inbound

Incentivizing High-Quality Human Annotations with Golden Questions cites this paper.

Incentivizing High-Quality Human Annotations with Golden Questions Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:42:19.330308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T13:41:26.730528Z digest=sha256:1b35e6a5bbc4c124efdec65413a79a8d735c215f1a5824170789a8ce39bb5638

Observation ab7a5c2b-8e98-45f3-8a3e-dcd03032f62b · inbound

On Symmetric Losses for Robust Policy Optimization with Noisy Preferences cites this paper.

On Symmetric Losses for Robust Policy Optimization with Noisy Preferences Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:54.719067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:54.719067Z digest=sha256:b834b7af26bd8639f720472174e5c005f1e6cae1e12bf70fae27716ad2cbfec4

Observation a3a8c523-4eb3-47c0-aee9-23fb7feeb443 · inbound

Influence Functions for Preference Dataset Pruning cites this paper.

Influence Functions for Preference Dataset Pruning Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:23.422447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:09:23.422447Z digest=sha256:bf9c17991d552203d0721b59b4f0226cf7684c4be718232d9ea4ce4ca3e4eeec

Observation 7a9371fc-802a-4a69-9862-70e064b6a897 · inbound

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap cites this paper.

Difficulty-Based Preference Data Selection by DPO Implicit Reward Gap Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:50:47.667921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T23:46:24.208438Z digest=sha256:befbd8a0699875f980a3a6fd5b120baa62f0dfb1bb7ff7338dab790e6bd76903

Observation 50ba8635-f3fe-402d-a776-48a192a2ef71 · inbound

Users as Annotators: LLM Preference Learning from Comparison Mode cites this paper.

Users as Annotators: LLM Preference Learning from Comparison Mode Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:21:06.939944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:19:58.093621Z digest=sha256:e1c0adf9d1c76b1dac4da2ef07866a0626f3300f941af45a7e5aba91d5d02e52

Observation 686cd752-1415-423f-b2a7-b5aea300d9d7 · inbound

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control cites this paper.

Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-22T11:21:29.105752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T11:17:03.104902Z digest=sha256:966f3f217524faf61e1f0d0ff58cd24da5f5b898465959b5e0c3dada593d525f

Observation 635b5e9b-43ee-4cbc-a20b-45bb1680db37 · inbound

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization cites this paper.

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:08.589391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T16:07:49.362916Z digest=sha256:c5e795f84d7f5dcc544f80ced0b8a37c90a3b01cc654feece380abb01d582750

Observation b10248ad-87c6-4251-969a-f78928ada1c9 · inbound

Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels cites this paper.

Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T15:37:01.391649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:37:01.391649Z digest=sha256:4fd7cb82e9673e84fc97306bef3fbe10f83df53843e83c2c63d0dbf8a2c9cf49

Observation 1c693273-2509-4423-9d5e-862d0b850da2 · inbound

Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels cites this paper.

Metadata-Free Meta-Reweighted Direct Preference Optimization under Noisy Preference Labels Impact of Preference Noise on the Alignment Performance of Generative Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T08:00:53.707181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:00:53.707181Z digest=sha256:ca18b521ee5a5227e12689e4368fdbacdd1cb3a30245e0a456d38c707ddbcf22