Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:49:12.618400Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 5 inbound Pith citation observations for arXiv:2505.23749.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:49:12.618400Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T22:50:33.210790Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T10:16:28.587559Z
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7deaf369-0115-498f-9e6f-35896dbed747 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? URL https://incidentdatabase.ai/
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f17035f9-7e11-455c-a298-7639db870ecf · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Statistical methods for ranking data, volume 1341
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f044d052-40b6-424c-b74d-4b9bf0b44fd6 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Approximating optimal social choice under metric preferences
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 39a3ad23-7acb-4a70-aaec-668482b00c9e · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Voudouris
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 281a5242-cd7a-4be5-aa31-39f775883da9 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? A general theoretical paradigm to understand learning from human preferences
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 933601e4-587f-4584-a255-27fc293941ac · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? A statistical decision-theoretic framework for social choice
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e2ad6e4-b9b9-4bbb-8bfb-0f2cabdaf924 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 124a96fa-0c60-489c-8166-cfaa13d61c76 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Constitutional AI: Harmlessness from AI Feedback
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d3a6849-baec-456b-9c58-9dea8b789c5b · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Procaccia, and Nisarg Shah
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a53979b3-731c-4b6b-9285-733542d78c26 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Procaccia, and Or Sheffet
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 135d9ea8-8faf-4c42-88a4-8e136d720380 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Human alignment of large language models through online preference optimisation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 74a561e3-88ef-4dfb-a4f0-09abee07dfa2 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Procaccia
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ceee655c-706a-4bed-9b54-bddb236a2c17 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? MaxMin-RLHF: Alignment with Diverse Human Preferences
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1505f2ee-acb4-4ea2-b434-c085a5c3cf15 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Breaking the Metric Voting Distortion Barrier
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 657b84e8-a411-47bf-b3d9-913419ba6b05 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b656a40b-4940-4c9c-9452-a3602d222e2b · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Chatbot arena: An open platform for evaluating llms by human preference
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bba5fc7f-0794-4c31-95f9-2e38d7b9a17d · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Direct preference optimization with unobserved preference heterogeneity
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3287c2d-66b1-4ba7-8719-2bad0b0111cd · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Deep reinforcement learning from human preferences
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f635759-00d9-4f10-a602-8e4f40f0544c · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Common voting rules as maximum likelihood estimators
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53035af2-5446-4811-8be4-e76807208457 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Position: social choice should guide ai alignment in dealing with diverse human feedback
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8255786f-6863-4923-b588-b353e24419ae · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Mapping social choice theory to RLHF
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f930d716-8fae-4698-a4a6-979d1ee73a3c · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Metric distortion with elicited pairwise comparisons
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f241ac89-4777-4b40-9923-727fd10053f3 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Optimized Distortion and Proportional Fairness in Voting
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ed957c40-ce0e-4abd-8715-ee0b64d970e2 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? KTO: Model Alignment as Prospect Theoretic Optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d7ac5d-9d6d-4b59-b981-cf5524b77920 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Fishburn
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 31fb267e-8df1-492f-8210-f97849efe89a · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Distortion Under Public-Spirited Voting
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4c3300b7-4f7e-47ce-81db-41a4f3551c64 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Axioms for AI Alignment from Human Feedback
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 443802cf-84bc-46c4-a923-a1cc58f2e459 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Resolving the optimal metric distortion conjecture
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d98f4e0b-79c1-45e8-af3a-b78a486267f6 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Metric distortion Under Probabilistic Voting
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b55ea79f-9f8e-422d-bb49-586c347b5e54 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 187a5caa-f303-427a-ab77-a8f27d4a248c · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Plurality Veto : A Simple Voting Rule Achieving Optimal Metric Distortion
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation af056ee5-f7c8-4972-b936-4e9d118cbfdf · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Provably mitigating overoptimization in RLHF : Your SFT loss is implicitly an adversarial regularizer
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 34e55da6-758a-4f5d-8540-f8d6d76bcb01 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Jackpot! Alignment as a Maximal Lottery
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 380fc395-09fe-4f14-86a5-85e5e85f348c · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? SimPO : Simple preference optimization with a reference-free reward
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eec2a541-f21b-427a-832c-569c9656b9d6 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? AI Alignment and Social Choice: Fundamental Limitations and Policy Implications
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cdd7021-dfb3-48b7-beed-bb8e171de671 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Nash learning from human feedback
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ec4ebb3b-c40d-4cb6-a5b3-2407794f695d · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Axioms for learning from pairwise comparisons
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d3e44c7f-c6df-4086-ba1e-66ccb2066cb7 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Training language models to follow instructions with human feedback
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3775ccd-ef83-4694-be14-831abd0f9ccf · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? RLHF from Heterogeneous Feedback via Personalization and Preference Aggregation
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dfc14ac-9814-4f7f-aaf3-284253afea04 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Personalizing reinforcement learning from human feedback with variational preference learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f0803d3b-e7d3-4d2e-b241-01c425f39311 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Procaccia and Jeffrey S
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f9158f3a-d4a9-4615-8655-ff963d98d703 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Clone-Robust AI Alignment
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a9b96989-7746-4bd0-a0a4-71dd9b38c846 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Direct preference optimization: Your language model is secretly a reward model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32346169-d22e-45e7-964e-082a9f96a582 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Proximal Policy Optimization Algorithms
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55c70479-b0f0-423f-9c06-9d3c7fb55913 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Direct Alignment with Heterogeneous Preferences
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d19d2f2-1c7b-4921-a9fd-08f0ea002ba9 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85be41e3-bb23-401c-9146-ccc3194f7a51 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Position: a roadmap to pluralistic alignment
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd7f7b4b-28b6-43e8-b9aa-3ae3317ff2c8 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? A Minimaximalist Approach to Reinforcement Learning from Human Feedback
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c9c893-73b3-4dc6-afdf-e447d80452e0 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Learning populations of preferences via pairwise comparison queries
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2389c91d-76bf-4b95-a76a-2daf0a194396 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Is RLHF more difficult than standard RL ? a theoretical perspective
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73729af5-8349-48e5-b23e-46786efbb058 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Metric learning from limited pairwise preference comparisons
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f4cbf904-d1a0-4470-bc7c-2ff692979b15 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Self-Play Preference Optimization for Language Model Alignment
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cd5d9d-8c3e-4ba5-b8b9-c87ae757715d · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Bayesian estimators as voting rules
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 95b897aa-fda4-43fc-9d1e-beac3deecb61 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Learning and decision-making from rank data
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4bdc156b-1a8c-44a1-927b-62721605be01 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? On the identifiability of mixtures of ranking models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2295f788-56e7-4694-a0bc-4acd973fa490 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Learning mixtures of plackett-luce models from structured partial orders
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a108ef03-f17e-45ac-a714-4404501048aa · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Learning mixtures of plackett-luce models
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86890663-deca-4538-9ae6-381b1a419079 · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23373edc-8627-46d1-a076-7e35ca23979c · outbound
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Fine-Tuning Language Models from Human Preferences
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2bb61c-f6a6-4a06-8322-dafc461ed06d · inbound
Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?
Reference 153
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 551547b7-8d69-4119-a622-2bb0160cf755 · inbound
Power and Limitations of Aggregation in Compound AI Systems Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f45ffd9c-70ac-4c7b-857c-6846556304ef · inbound
Mind the Gap: Structure-Aware Consistency in Preference Learning Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1da2f02c-27f8-440b-8fa5-c2ad82146aba · inbound
Internal Pluralism and the Limits of Pairwise Comparisons Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?
Reference 133
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0da2bd54-9479-4abd-a04d-c364a1e6eb54 · inbound
Internal Pluralism and the Limits of Pairwise Comparisons Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences?
Reference 222
Source-reported events for the cited work
Unavailable: canonical work link unavailable.