REVIEW 2 major objections 5 minor 296 references
Internal Pluralism and the Limits of Pairwise Comparisons
T0 review · 2 major / 5 minor · reviewed 2026-08-02 · deepseek-v4-flash
Pith's one-line read Local pairwise comparisons—the standard tool for learning what people want—cannot detect global priorities like egalitarianism, and can be systematically distorted by internal conflict; allowing people to report indecision substantially acc
desk verdict A clean formal model showing local pairwise comparisons can miss global priorities, but its central erasure theorem rests on an untested assumption about how people aggregate across background rules. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing object is the pluralistic priority model: a weighted collection of 'priorities', each a complete, transitive preference relation over the space of decision rules, together with a two-stage aggregation. For a local query (y,y';x), each priority first aggregates its 'projection gaps' across all background rules via a permutation-invariant, unanimous rule aggregator; the priorities' weighted evidence is then tallied into two directional scores s+ and s-, whose sum and difference give valence and decisiveness. Thresholds on these quantities produce four latent states: decisive preference either way, indifference (too little total evidence), and conflict (strong evidence both way
What would settle it
Take the two-recipient, two-input allocation example from the paper (Proposition 3.2), where egalitarianism is perfectly inseparable. Ask participants a set of local pairwise comparisons, fit a Bradley-Terry model, and then ask them to choose between complete decision rules (e.g., a rule that alternates between the two recipients vs. a rule that always gives to one). If participants' global rule choices are predicted by the local-fit model, the erasure phenomenon is behaviorally irrelevant; if the global choices reveal a preference for balanced rules that the local comparisons could not have p
Extended reading notes
Core claim
On the paper's own terms, the central discovery is Theorem 3.3: any model that contains a nonzero number of perfectly inseparable priorities is behaviorally indistinguishable, over local pairwise comparisons, from the same model with those priorities deleted. Because a permutation-invariant rule aggregator sees perfectly inseparable priorities as contributing identical evidence for both options at every query, their weight cancels from the decisiveness signal up to a scale factor that is absorbed by the noise temperature. Consequently, any decoder that fits a perfectly separable model (the class that includes Bradley-Terry) will erase these priorities without a trace on exhaustive data (Coro
Load-bearing premise
The central claim depends on the assumption that, when answering a local question, a person summarizes each priority by aggregating its implications across all possible 'background' behaviors of the rule symmetrically (permutation-invariantly), before combining priorities—if instead people weight some background scenarios as more plausible or salient, perfectly inseparable priorities would not cancel and could be learned from local comparisons.
Editorial extensions
If this is right
- Standard preference-learning pipelines built on Bradley-Terry or other score-based random utility models are, by Theorem 2.3, restricted to perfectly separable priorities with no indecision; the paper's results therefore delimit exactly when those pipelines are justified.
- Perfectly inseparable priorities—formalized versions of egalitarianism, proportionality, and equal treatment—can be erased without a trace, so a learner may converge to a rule that is near-worst according to the individual's true priorities even with unlimited data.
- Because perfectly inseparable weights are non-identifiable from local comparisons, no amount of additional local data can recover them; richer query formats are necessary.
- Allowing people to report indecision (even a single generic option conflating conflict and indifference) substantially reduces the number of queries needed to learn accurate priority weights, with gains concentrated in the tens-of-queries regime that is practical for elicitation.
- Learning weights is not the same as learning rules: in the forced-response simulations, weight errors were large (14–24% of worst case) while average regret was modest, but worst-case regret reached up to 39% of the utility range.
Reading between the lines
- If local comparisons cannot identify inseparable priorities, then eliciting priorities directly—for example, asking people to state their criteria in text and then weighting those criteria—becomes not just a convenience but a necessary complement to pairwise data; the paper sketches this 'priority-aware learning' direction.
- The rule-first aggregation assumption (aggregate evidence over background rules before combining priorities) is what makes the perfect-inseparability cancellation go through; if individuals instead weight background rules by plausibility or context, inseparable priorities would leave a detectable trace. Testing this aggregation order behaviorally is a natural next step.
- The paper's formal distinction between indifference and conflict maps onto behavioral findings that people defer choice under conflict but not under indifference; a testable extension would compare learning speed when the indecision option is framed as 'no strong feeling' versus 'genuinely torn'.
- Because the indecision benefit persists even when the learner does not know the response thresholds (learned jointly), the acceleration is not an artifact of privileged information; this suggests practical systems can offer an indecision button without a separate calibration phase.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper proposes a formal model of "internal pluralism" in which an individual's preferences over decision rules are represented by multiple weighted priorities, each a complete transitive utility over the rule space. It defines how these rule-level priorities generate local pairwise comparison responses through two aggregation steps: first over background rules within each priority, then over priorities, yielding directional evidence scores and latent states of decisiveness, conflict, and indifference. The contributions are: (i) a characterization of score-based random utility models (S-RUMs), including Bradley-Terry, as exactly the perfectly separable, zero-threshold special case of this model (Theorem 2.3); (ii) results showing that perfectly inseparable priorities—such as certain formalizations of egalitarianism, proportionality, and equal treatment—are erased from local pairwise comparison data and are not identifiable, while general inseparable priorities can be misinterpreted by separable-consistent decoders (Theorem 3.3, Corollaries 3.4–3.5, Example 4); and (iii) simulations in linear priority models showing that forced decisive responses under latent indecision cause weight-estimation and worst-case regret losses, and that explicitly allowing indecision reports substantially accelerates Bayesian active learning. The paper concludes by proposing "priority-aware learning" as a more faithful elicitation paradigm.
Significance. If the results are taken at face value, this is a significant conceptual contribution to preference learning and AI alignment. The formal equivalence between S-RUMs and the zero-threshold separable case is clean and useful, and the non-identification theorem for perfectly inseparable priorities is a genuine formal obstruction to a widely used elicitation pipeline. The simulation study is carefully designed, with generative data, hyperparameter settings, multiple response models, standard errors, and an explicit active-learning procedure; the finding that indecision reports contain useful information beyond avoiding forced responses is plausible and practically relevant. The paper is also commendable for including detailed appendix proofs, explicit assumptions, and a candid discussion of limitations. However, the breadth of the title and abstract exceeds what the formal model actually establishes, because the central erasure and non-identification results are contingent on a specific, behaviorally unvalidated aggregation architecture. The paper is therefore more a rigorous conditional result than a general impossibility theorem about local pairwise comparisons.
major comments (2)
- The core non-identification result is derived under the assumption that aggregation happens over background rules first, then over priorities, with a permutation-invariant and unanimous rule aggregator. The authors explicitly defer other aggregation orders to future work, but the manuscript's framing—"erased without a trace" and "not identifiable by local pairwise comparisons at all"—presents this as a property of local pairwise comparisons themselves, not of a particular model architecture. Under a plausible alternative, priority-first aggregation with a max aggregator, a perfectly inseparable egalitarian priority does not cancel: the decisiveness score kappa(q) can depend on the egalitarian weight, so the priority leaves a detectable trace and the conclusion of Theorem 3.3 fails. Since no behavioral evidence is offered for rule-first aggregation, the load-bearing claim is conditional o
- The general-inseparability example relies on an additional "balanced sign-responsive" condition on the rule aggregator, but this condition is not stated in the main text and is not satisfied by all aggregators permitted by Definition 2.3 (it excludes, for example, lower-percentile aggregators). The formal construction in Appendix B.6 correctly proves the claim under this extra assumption, but the main-text Example 4 reads as a general demonstration of misinterpretation of inseparable priorities. This is a load-bearing point for the paper's broader narrative that general inseparable priorities are systematically distorted by separable-consistent decoders. The condition should be stated in the main text, and the scope of the claim should be narrowed to aggregators satisfying it, or the authors should explain why the restriction is harmless.
minor comments (5)
- Typo: "assumpions" should be "assumptions".
- Heading "Eqal Treatment" should be "Equal Treatment".
- The quantifier structure in scale-free indistinguishability is ambiguous: it should be made explicit whether the rescaling beta' is allowed to depend on the query q or must be uniform across all queries. Lemma B.4 requires a uniform rescaling, and the proof of Theorem 3.3 constructs one, but the definition as written can be read otherwise.
- The paper shows that egalitarianism can be perfectly inseparable, but only in a two-recipient, two-input toy example. The abstract's broad statement that "egalitarianism ... local comparisons may fail to capture them" is supported by the weaker inseparability result (Proposition 3.1), but the strong erasure result (Corollary 3.4) applies only to perfect inseparability. A sentence clarifying the gap between the general notion and the toy instances would help readers avoid overgeneralizing.
- The simulation treats inputs as drawn from a continuous distribution despite the model's finite-X assumption; the authors note this and suggest discretization, but a more careful statement of how the reported numbers map to the formal model would improve reproducibility.
Circularity Check
No significant circularity: the central results unpack explicit definitions with self-contained proofs, and the simulations are generative rather than fitted to targets.
full rationale
The paper's central formal results are derived from explicit definitions rather than importing the conclusion. Theorem 2.3 constructively maps S-RUMs to the perfectly separable zero-threshold case in both directions, and the optimal-rule correspondence is proven by a telescoping argument, so the equivalence is not assumed. Theorem 3.3 and its corollaries follow directly from Definition 3.1 (perfect inseparability as an equal-size symmetric evidence profile) and the stated permutation-invariance of the rule aggregator; the cancellation is a proof from those definitions, not a fitted parameter renamed as a prediction. The non-trivial content is carried by Propositions 3.1, 3.2, B.1, and B.2, which show that natural priorities such as egalitarianism, proportionality, and equal treatment can satisfy the strong definition in non-degenerate constructed examples. The Section 4 experiments are generative: the learner is not fit to observed indecision labels, and the result that indecision reports accelerate learning is a property of the defined response model. The paper explicitly acknowledges the favorable assumption that individuals report indecision accurately, which is a limitation rather than a circular step. The rule-first aggregation assumption is an unvalidated modeling choice that affects external validity, but it is not a circular reduction: the proofs do not assume the non-identifiability or distortion conclusions. The only self-citation (Baharav et al. 2026, used to contextualize cosine similarity and regret) is not load-bearing. Overall, no load-bearing step is equivalent to its own input by construction, so there is no significant circularity.
Assumptions & free parameters
free parameters (5)
- Thresholds τr, τκ =
τr=τκ=0.25 for §4.2; grid {0, 0.2, ..., 0.8} for §4.3
- Logistic inverse temperature β =
10
- Dirichlet concentration for true weights =
0.2
- Example perturbation ε =
0 < ε < 1/4 in Examples 3–4
- BALD/hyperparameters (C=50, Npost=200, NBALD=50, burn-in=200) =
See Table 2
assumptions (6)
- domain assumption Each priority is a complete, transitive weak preference relation over rules with a utility representation normalized to bounded gaps
- ad hoc to paper Aggregation happens over rules first, then over priorities; φrules is permutation-invariant and unanimous; priority aggregators are scale-preserving
- standard math Link function h is continuous, strictly increasing, and odd-symmetric (logit/probit form)
- domain assumption In Section 4, priorities are linear in a known feature space and the feature map is known to the learner
- ad hoc to paper In Section 4.2, individuals deviate from the learner's model only on latently indecisive queries
- ad hoc to paper Perfect inseparability is defined with equal-size partitions and equal δ
invented entities (3)
-
Priority model M=(u,ω)
-
Latent states of conflict (⊳) vs indifference (~)
-
Perfect (in)separability classification
Cite this review
Pith. "Pith review of Internal Pluralism and the Limits of Pairwise Comparisons." pith.science (2026). https://pith.science/paper/RABX723P
@misc{pith2026260702672,
author = {Pith},
title = {Pith review of: Internal Pluralism and the Limits of Pairwise Comparisons},
year = {2026},
howpublished = {\url{https://pith.science/paper/RABX723P}},
note = {Machine review of arXiv:2607.02672}
}
read the original abstract
Local pairwise comparisons are a standard tool for learning how people want decision rules to work, e.g., in participatory design or alignment. However, their use builds in two strong assumptions: that local comparisons are sufficient evidence about how a person wants an automated decision rule to behave, and that people can always answer those comparisons decisively. We investigate how these assumptions may be compromised under internal pluralism: the idea that an individual evaluates decision rules according to multiple authoritative priorities about how the rule should behave. We provide a formal model of such pluralistic preferences over decision rules, which then lets us identify two distinct failures of forced local pairwise comparison data. First, priorities such as proportionality, egalitarianism, and equal treatment are inherently global: what they imply in one case can depend on what happens elsewhere, so local comparisons may fail to capture them. Second, even when priorities are representable locally, tension between strongly-held priorities can generate internal conflict, producing potentially costly behavioral distortions when comparisons are forced. We then use our model to investigate the alternative -- allowing people to report indecision -- and our findings suggest that doing so can considerably reduce the number of queries needed to learn preferences accurately. We conclude by describing how our model points toward preference-learning methods that elicit these priorities directly, yielding more faithful and interpretable accounts of what people value.
Figures
Figures from the paper (12 more)
Reference graph
Works this paper leans on
-
[1]
arXiv preprint arXiv:2601.13178 , year=
Medical Triage as Pairwise Ranking: A Benchmark for Urgency in Patient Portal Messages , author=. arXiv preprint arXiv:2601.13178 , year=
-
[2]
arXiv preprint arXiv:2603.19510 , year=
Linear Social Choice with Few Queries: A Moment-Based Approach , author=. arXiv preprint arXiv:2603.19510 , year=
-
[3]
and Kehne, G
Ge, L. and Kehne, G. and Vorobeychik, Y. , title =. Proceedings of the AAAI Conference on Artificial Intelligence , volume =. 2026 , doi =
2026
-
[4]
Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Maple: A framework for active preference learning guided by large language models , author=. Proceedings of the AAAI Conference on Artificial Intelligence , volume=
-
[5]
, title =
Feffer, Michael and Heidari, Hoda and Lipton, Zachary C. , title =. Proceedings of the 37th AAAI Conference on Artificial Intelligence (AAAI) , pages =
-
[6]
arXiv preprint arXiv:2605.12717 , year=
The End Justifies the Mean: A Linear Ranking Rule for Proportional Sequential Decisions , author=. arXiv preprint arXiv:2605.12717 , year=
-
[7]
Psychological Review , volume =
Amos Tversky , title =. Psychological Review , volume =
-
[8]
Efficient Pairwise Preference Elicitation Allowing for Indifference , journal =
J. Efficient Pairwise Preference Elicitation Allowing for Indifference , journal =
Show all 296 references
-
[9]
and Nelson, Barry L
Cario, Marne C. and Nelson, Barry L. , title =. Technical Report, Department of Industrial Engineering and Management Sciences, Northwestern University , year =
-
[10]
arXiv preprint arXiv:2604.03238 , year =
Bijean Ghafouri and Eun Cheol Choi and Priyanka Dey and Emilio Ferrara , title =. arXiv preprint arXiv:2604.03238 , year =
-
[11]
arXiv preprint arXiv:2310.11589 , year=
Eliciting human preferences with language models , author=. arXiv preprint arXiv:2310.11589 , year=
-
[12]
1980 , publisher =
The Analytic Hierarchy Process , author =. 1980 , publisher =
1980
-
[13]
Krosnick , title =
Jon A. Krosnick , title =. Survey Nonresponse , editor =
-
[14]
Voting by Committees , journal =
Barber. Voting by Committees , journal =. 1991 , doi =
1991
-
[15]
Voting by Committees under Constraints , journal =
Barber. Voting by Committees under Constraints , journal =. 2005 , doi =
2005
-
[16]
Voting in Combinatorial Domains , booktitle =
Lang, J. Voting in Combinatorial Domains , booktitle =. 2016 , doi =
2016
-
[17]
Sequential Composition of Voting Rules in Multi-Issue Domains , journal =
Lang, J. Sequential Composition of Voting Rules in Multi-Issue Domains , journal =. 2009 , doi =
2009
-
[18]
Conditional and Sequential Approval Voting on Combinatorial Domains , booktitle =
Barrot, Nathana. Conditional and Sequential Approval Voting on Combinatorial Domains , booktitle =
-
[19]
, title =
Ratliff, Thomas C. , title =. Public Choice , volume =. 2006 , doi =
2006
-
[20]
and Saari, Donald G
Ratliff, Thomas C. and Saari, Donald G. , title =. Social Choice and Welfare , volume =. 2014 , doi =
2014
-
[21]
ECAI 2010: 19th European Conference on Artificial Intelligence , editor =
Uckelman, Joel , title =. ECAI 2010: 19th European Conference on Artificial Intelligence , editor =. 2010 , doi =
2010
-
[22]
2026 , eprint =
Ge, Luise and Halpern, Daniel and Kehne, Gregory and Vorobeychik, Yevgeniy , title =. 2026 , eprint =
2026
-
[23]
1993 , publisher=
Decisions with multiple objectives: preferences and value trade-offs , author=. 1993 , publisher=
1993
-
[24]
4OR-A Quarterly Journal of Operations Research , number=
Preference learning and multiple criteria decision aiding: differences, commonalities, and synergies—part II , author=. 4OR-A Quarterly Journal of Operations Research , number=. 2024 , publisher=
2024
-
[25]
International Conference on Learning Representations , volume=
Distributional preference learning: Understanding and accounting for hidden context in rlhf , author=. International Conference on Learning Representations , volume=
-
[26]
2020 , publisher=
Moral uncertainty , author=. 2020 , publisher=
2020
-
[27]
Journal of Logic and Computation , volume=
Persuasion in practical argument using value-based argumentation frameworks , author=. Journal of Logic and Computation , volume=. 2003 , publisher=
2003
-
[28]
, title =
Fehr, Ernst and Schmidt, Klaus M. , title =. The Quarterly Journal of Economics , volume =
-
[29]
and Hole, Astri Drange and S
Cappelen, Alexander W. and Hole, Astri Drange and S. The Pluralism of Fairness Ideals: An Experimental Approach , journal =
-
[30]
American Economic Review , volume =
Konow, James , title =. American Economic Review , volume =
-
[31]
, title =
Leventhal, Gerald S. , title =. Social Exchange: Advances in Theory and Research , editor =
-
[32]
European Journal of Operational Research , year =
The Analytic Hierarchy Process in Medical and Health Care Decision Making: A Literature Review , author =. European Journal of Operational Research , year =
-
[33]
Frontiers in Econometrics , year =
Conditional Logit Analysis of Qualitative Choice Behavior , author =. Frontiers in Econometrics , year =
-
[34]
1985 , publisher=
Discrete choice analysis: theory and application to travel demand , author=. 1985 , publisher=
1985
-
[35]
2000 , publisher =
Stated Choice Methods: Analysis and Applications , author =. 2000 , publisher =
2000
-
[36]
Quality in Health Care , year =
Using Discrete Choice Experiments to Elicit Preferences , author =. Quality in Health Care , year =
-
[37]
Advances in Neural Information Processing Systems , year =
A Bayesian Approach for Policy Learning from Trajectory Preference Queries , author =. Advances in Neural Information Processing Systems , year =
-
[38]
Robotics: Science and Systems (RSS) , year =
Active Preference-Based Learning of Reward Functions , author =. Robotics: Science and Systems (RSS) , year =
-
[39]
Science , year =
The Social Dilemma of Autonomous Vehicles , author =. Science , year =
-
[40]
Proceedings of the 17th International Conference on Machine Learning (ICML) , year =
Algorithms for Inverse Reinforcement Learning , author =. Proceedings of the 17th International Conference on Machine Learning (ICML) , year =
-
[41]
Proceedings of the 21st International Conference on Machine Learning (ICML) , year =
Apprenticeship Learning via Inverse Reinforcement Learning , author =. Proceedings of the 21st International Conference on Machine Learning (ICML) , year =
-
[42]
AAAI Conference on Artificial Intelligence , year =
Maximum Entropy Inverse Reinforcement Learning , author =. AAAI Conference on Artificial Intelligence , year =
-
[43]
Journal of Machine Learning Research , year =
Classification with a Reject Option using a Hinge Loss , author =. Journal of Machine Learning Research , year =
-
[44]
Journal of Machine Learning Research , year =
On the Foundations of Noise-free Selective Classification , author =. Journal of Machine Learning Research , year =
-
[45]
, title =
Graham, Jesse and Haidt, Jonathan and Nosek, Brian A. , title =. Journal of Personality and Social Psychology , year =
-
[46]
and Haidt, Jonathan and Iyer, Ravi and Koleva, Spassena and Ditto, Peter H
Graham, Jesse and Nosek, Brian A. and Haidt, Jonathan and Iyer, Ravi and Koleva, Spassena and Ditto, Peter H. , title =. Journal of Personality and Social Psychology , year =
-
[47]
Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems , pages=
Expressiveness, Cost, and Collectivism: How the Design of Preference Languages Shapes Participation in Algorithmic Decision-Making , author=. Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems , pages=
2023
-
[48]
The least squares solution assuming equal standard deviations and equal correlations , author=
Remarks on the method of paired comparisons: I. The least squares solution assuming equal standard deviations and equal correlations , author=. Psychometrika , volume=. 1951 , publisher=
1951
-
[49]
2023 , publisher=
Multi-winner voting with approval preferences , author=. 2023 , publisher=
2023
-
[50]
and Kim, Jae Il , title =
Tetlock, Philip E. and Kim, Jae Il , title =. Journal of Personality and Social Psychology , year =
-
[51]
Psychological Review , year =
Haidt, Jonathan , title =. Psychological Review , year =
-
[52]
Moral Dumbfounding: When Intuition Finds No Reason , institution =
Haidt, Jonathan and Bj. Moral Dumbfounding: When Intuition Finds No Reason , institution =. 2000 , note =
2000
-
[53]
Consequences, Norms, and Generalized Inaction in Moral Dilemmas: The
Gawronski, Bertram and Armstrong, Joel and Conway, Paul and Friesdorf, Rebecca and H. Consequences, Norms, and Generalized Inaction in Moral Dilemmas: The. Journal of Personality and Social Psychology , volume =. 2017 , doi =
2017
-
[54]
and Luke, Dillon M
Ng, Nyx L. and Luke, Dillon M. and Gawronski, Bertram , title =. Personality and Social Psychology Bulletin , year =. doi:10.1177/01461672231180760 , note =
-
[55]
2019 , publisher =
Human Compatible: Artificial Intelligence and the Problem of Control , author =. 2019 , publisher =
2019
-
[56]
Trends in cognitive sciences , volume=
Thinking the unthinkable: Sacred values and taboo cognitions , author=. Trends in cognitive sciences , volume=. 2003 , publisher=
2003
-
[57]
2nd Workshop on Models of Human Feedback for AI Alignment , year=
Full-stack alignment: Co-aligning AI and institutions with thicker models of value , author=. 2nd Workshop on Models of Human Feedback for AI Alignment , year=
-
[58]
arXiv preprint arXiv:2112.00861 , year=
A general language assistant as a laboratory for alignment , author=. arXiv preprint arXiv:2112.00861 , year=
-
[59]
Advances in Neural Information Processing Systems , volume=
Clave: An adaptive framework for evaluating values of llm generated responses , author=. Advances in Neural Information Processing Systems , volume=
-
[60]
Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) , pages=
Value FULCRA: Mapping large language models to the multidimensional spectrum of basic human value , author=. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) , pages=
2024
-
[61]
arXiv preprint arXiv:2505.12792 , year=
EAVIT: Efficient and Accurate Human Value Identification from Text data via LLMs , author=. arXiv preprint arXiv:2505.12792 , year=
-
[62]
and Griffith, H
Thapa, B. and Griffith, H. and Rathore, H. , title =. Security and Privacy in Cyber-Physical Systems and Smart Vehicles , editor =. 2025 , note =. doi:10.1007/978-3-031-93354-7_16 , url =
2025 doi
-
[63]
Proceedings of the AAAI Conference on Artificial Intelligence , author =
Participatory. Proceedings of the AAAI Conference on Artificial Intelligence , author =. 2023 , note =. doi:10.1609/aaai.v37i5.25699 , abstract =
2023 doi
-
[64]
Transportation Research Part F: Traffic Psychology and Behaviour , volume=
Decision modeling for automated driving in dilemmas based on bidirectional value alignment of moral theory values and fair human moral values , author=. Transportation Research Part F: Traffic Psychology and Behaviour , volume=. 2025 , publisher=
2025
-
[65]
1998 , publisher =
Asymptotic Statistics , author =. 1998 , publisher =
1998
-
[66]
The Annals of Mathematical Statistics , volume =
Limiting Behavior of Posterior Distributions when the Model is Incorrect , author =. The Annals of Mathematical Statistics , volume =. 1966 , publisher =
1966
-
[67]
The Annals of Statistics , volume =
Misspecification in Infinite-Dimensional Bayesian Statistics , author =. The Annals of Statistics , volume =. 2006 , publisher =
2006
-
[68]
Schervish , keywords =
Taeryon Choi and Mark J. Schervish , keywords =. On posterior consistency in nonparametric regression problems , journal =. 2007 , issn =. doi:https://doi.org/10.1016/j.jmva.2007.01.004 , url =
2007 doi
-
[69]
Journal of Statistical Software , author =
Bradley-. Journal of Statistical Software , author =. 2012 , pages =. doi:10.18637/jss.v048.i09 , abstract =
2012 doi
-
[70]
Political Analysis , author=
Causal Inference in Conjoint Analysis: Understanding Multidimensional Choices via Stated Preference Experiments , volume=. Political Analysis , author=. 2014 , pages=. doi:10.1093/pan/mpt024 , number=
2014 doi
-
[71]
The Limitations of Using Forced Choice in Electoral Conjoint Experiments , author=
-
[72]
ERC: A Theory of Equity, Reciprocity, and Competition , urldate =
Gary E Bolton and Axel Ockenfels , journal =. ERC: A Theory of Equity, Reciprocity, and Competition , urldate =
-
[73]
Proceedings of the National Academy of Sciences , volume=
A moral trade-off system produces intuitive judgments that are rational and coherent and strike a balance between conflicting moral values , author=. Proceedings of the National Academy of Sciences , volume=. 2022 , publisher=
2022
-
[74]
and Raiffa, Howard , title =
Keeney, Ralph L. and Raiffa, Howard , title =
-
[75]
Frontiers in Econometrics , editor =
McFadden, Daniel , title =. Frontiers in Econometrics , editor =. 1974 , pages =
1974
-
[76]
and Rao, Vithala R
Green, Paul E. and Rao, Vithala R. , title =. Journal of Marketing Research , year =
-
[77]
and Sommerville, R
Greene, Joshua D. and Sommerville, R. Brian and Nystrom, Leigh E. and Darley, John M. and Cohen, Jonathan D. , journal =. An. 2001 , doi =
2001
-
[78]
Organizational Behavior and Human Decision Processes , year =
Baron, Jonathan and Spranca, Mark , title =. Organizational Behavior and Human Decision Processes , year =
-
[79]
Proceedings of the National Academy of Sciences , year =
Danziger, Shai and Levav, Jonathan and Avnaim-Pesso, Liora , title =. Proceedings of the National Academy of Sciences , year =
-
[80]
Tory , title =
Higgins, E. Tory , title =. Social Psychology: Handbook of Basic Principles , editor =. 1996 , pages =
1996
-
[81]
and Chen, Mark and Burrows, Lara , title =
Bargh, John A. and Chen, Mark and Burrows, Lara , title =. Journal of Personality and Social Psychology , year =
-
[82]
and Cacioppo, John T
Petty, Richard E. and Cacioppo, John T. , title =. 1986 , doi =
1986
-
[83]
and Luskin, Robert C
Fishkin, James S. and Luskin, Robert C. , title =. Acta Politica , year =
-
[84]
arXiv preprint arXiv:1706.03741 , year=
Deep reinforcement learning from human preferences , author=. arXiv preprint arXiv:1706.03741 , year=
-
[86]
Advances in Neural Information Processing Systems (NeurIPS) , year=
Learning to summarize from human feedback , author=. Advances in Neural Information Processing Systems (NeurIPS) , year=
-
[87]
Advances in Neural Information Processing Systems (NeurIPS) , year=
Training language models to follow instructions with human feedback , author=. Advances in Neural Information Processing Systems (NeurIPS) , year=
-
[88]
arXiv preprint arXiv:2204.05862 , year=
Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback , author=. arXiv preprint arXiv:2204.05862 , year=
-
[90]
Journal of Machine Learning Research , volume=
A survey of preference-based reinforcement learning methods , author=. Journal of Machine Learning Research , volume=
-
[91]
Pharmacoeconomics , volume=
Conducting discrete choice experiments to inform healthcare decision making: a user’s guide , author=. Pharmacoeconomics , volume=. 2008 , publisher=
2008
-
[92]
Journal of Consumer Research , volume=
Conjoint Analysis in Consumer Research: Issues and Outlook , author=. Journal of Consumer Research , volume=
-
[93]
Journal of Marketing , volume=
Conjoint Analysis in Marketing: New Developments with Implications for Research and Practice , author=. Journal of Marketing , volume=
-
[94]
Frontiers in Econometrics , editor=
Conditional Logit Analysis of Qualitative Choice Behavior , author=. Frontiers in Econometrics , editor=
-
[95]
Robotics: Science and Systems (RSS) , year=
Active Preference-Based Learning of Reward Functions , author=. Robotics: Science and Systems (RSS) , year=
-
[96]
arXiv preprint arXiv:1810.04303 , year=
Batch Active Preference-Based Learning of Reward Functions , author=. arXiv preprint arXiv:1810.04303 , year=
-
[97]
arXiv preprint arXiv:2405.14758 , year=
Axioms for AI Alignment from Human Feedback , author=. arXiv preprint arXiv:2405.14758 , year=
-
[98]
Advances in Neural Information Processing Systems , volume =
Cooperative Inverse Reinforcement Learning , author =. Advances in Neural Information Processing Systems , volume =
-
[99]
Advances in Neural Information Processing Systems (NeurIPS) , year=
Policy Aggregation , author=. Advances in Neural Information Processing Systems (NeurIPS) , year=
-
[100]
arXiv preprint arXiv:2501.09254 , year=
Clone-Robust AI Alignment , author=. arXiv preprint arXiv:2501.09254 , year=
-
[101]
arXiv preprint arXiv:2502.16320 , year=
Direct Alignment with Heterogeneous Preferences , author=. arXiv preprint arXiv:2502.16320 , year=
-
[102]
International Conference on Machine Learning (ICML) , year=
Provably Efficient Preference-Based Reinforcement Learning with General Function Approximation , author=. International Conference on Machine Learning (ICML) , year=
-
[103]
Advances in Neural Information Processing Systems (NeurIPS) , year=
Reward-rational (implicit) choice: A unifying formalism for reward learning , author=. Advances in Neural Information Processing Systems (NeurIPS) , year=
-
[104]
arXiv preprint arXiv:1809.03060 , year=
Active Inverse Reward Design , author=. arXiv preprint arXiv:1809.03060 , year=
-
[105]
Applied Health Economics and Health Policy , year=
Using discrete choice experiments to value health care programmes: current practice and future research reflections , author=. Applied Health Economics and Health Policy , year=
-
[106]
Stated Choice Methods: Analysis and Applications , author=
-
[107]
Discrete Choice Methods with Simulation , author=
-
[108]
Political Analysis , volume=
Causal Inference in Conjoint Analysis: Understanding Multidimensional Choices via Stated Preference Experiments , author=. Political Analysis , volume=
-
[109]
arXiv preprint arXiv:2212.08073 , year=
Constitutional ai: Harmlessness from ai feedback , author=. arXiv preprint arXiv:2212.08073 , year=
-
[110]
Advances in neural information processing systems , volume=
Inverse reward design , author=. Advances in neural information processing systems , volume=
-
[111]
Rao, P. V. and Kupper, L. L. , title =. Journal of the American Statistical Association , volume =. 1967 , doi =
1967
-
[112]
Davidson , journal =
Roger R. Davidson , journal =. On Extending the Bradley-Terry Model to Accommodate Ties in Paired Comparison Experiments , urldate =
-
[113]
Schmidt , journal =
Ernst Fehr and Klaus M. Schmidt , journal =. A Theory of Fairness, Competition, and Cooperation , urldate =
-
[114]
, title =
Atkinson, Anthony B. , title =. Journal of Economic Theory , volume =. 1970 , issn =. doi:10.1016/0022-0531(70)90039-6 , url =
1970 doi
-
[115]
Preference for Flexibility
David M. Kreps , journal =. A Representation Theorem for "Preference for Flexibility" , urldate =
-
[116]
Aumann , journal =
Robert J. Aumann , journal =. Utility Theory without the Completeness Axiom , urldate =
-
[117]
American Journal of Political Science , year=
Are Americans Ambivalent Towards Racial Policies , author=. American Journal of Political Science , year=
-
[118]
arXiv preprint arXiv:2003.01899 , year=
Robust active preference elicitation , author=. arXiv preprint arXiv:2003.01899 , year=
2003 arXiv
-
[119]
Proceedings of the AAAI Conference on Artificial Intelligence , volume=
A voting-based system for ethical decision making , author=. Proceedings of the AAAI Conference on Artificial Intelligence , volume=
-
[120]
Nature , volume=
The moral machine experiment , author=. Nature , volume=. 2018 , publisher=
2018
-
[121]
Proceedings of the AAAI Conference on Artificial Intelligence , author =
Proportionally. Proceedings of the AAAI Conference on Artificial Intelligence , author =. 2021 , note =. doi:10.1609/aaai.v35i6.16646 , abstract =
2021 doi
-
[122]
Journal of Environmental Economics and Management , author =
How. Journal of Environmental Economics and Management , author =. 1994 , pages =. doi:10.1006/jeem.1994.1006 , abstract =
1994
-
[123]
2025 , eprint=
A Framework for Human-Reason-Aligned Trajectory Evaluation in Automated Vehicles , author=. 2025 , eprint=
2025
-
[124]
Proceedings of the 2024 ACM Conference on Fairness, Accountability, and Transparency , pages=
Participatory Objective Design via Preference Elicitation , author=. Proceedings of the 2024 ACM Conference on Fairness, Accountability, and Transparency , pages=
2024
-
[125]
Artificial Intelligence , volume=
Adapting a kidney exchange algorithm to align with human values , author=. Artificial Intelligence , volume=. 2020 , publisher=
2020
-
[126]
, author=
Elimination by aspects: A theory of choice. , author=. Psychological review , volume=. 1972 , publisher=
1972
-
[127]
, author=
Reasoning the fast and frugal way: models of bounded rationality. , author=. Psychological review , volume=. 1996 , publisher=
1996
-
[128]
, author=
Linear models in decision making. , author=. Psychological bulletin , volume=. 1974 , publisher=
1974
-
[129]
Management science , volume=
The challenge of understanding what users want: Inconsistent preferences and engagement optimization , author=. Management science , volume=. 2024 , publisher=
2024
-
[130]
Cognition , year =
Shafir, Eldar and Simonson, Itamar and Tversky, Amos , title =. Cognition , year =
-
[131]
Journal of Consumer Psychology , volume=
Will I like a “medium” pillow? Another look at constructed and inherent preferences , author=. Journal of Consumer Psychology , volume=. 2008 , publisher=
2008
-
[132]
Daedalus , volume=
Intuitive ethics: How innately prepared intuitions generate culturally variable virtues , author=. Daedalus , volume=. 2004 , publisher=
2004
-
[133]
Journal of Personality and Social Psychology , volume=
A value pluralism model of ideological reasoning , author=. Journal of Personality and Social Psychology , volume=. 1986 , publisher=
1986
-
[134]
Philosophy & Public Affairs , volume=
Rational fools: A critique of the behavioral foundations of economic theory , author=. Philosophy & Public Affairs , volume=. 1977 , publisher=
1977
-
[135]
Theory and Decision , volume=
Metapreferences and the reasons for stability in social choice: Thoughts on analyzing the dynamics of political conflict , author=. Theory and Decision , volume=. 1985 , publisher=
1985
-
[136]
1993 , publisher=
The Adaptive Decision Maker , author=. 1993 , publisher=
1993
-
[137]
2001 , publisher=
Bounded rationality: The adaptive toolbox , author=. 2001 , publisher=
2001
-
[138]
The American Economic Review , volume=
The causes of preference reversal , author=. The American Economic Review , volume=. 1990 , publisher=
1990
-
[139]
American Psychologist , volume=
The construction of preference , author=. American Psychologist , volume=. 1995 , publisher=
1995
-
[140]
Journal of Experimental Psychology: Learning, Memory, and Cognition , volume=
Aspects of endowment: a query theory of value construction , author=. Journal of Experimental Psychology: Learning, Memory, and Cognition , volume=. 2007 , publisher=
2007
-
[141]
Annual Review of Psychology , volume=
Mindful judgment and decision making , author=. Annual Review of Psychology , volume=. 2009 , publisher=
2009
-
[142]
Political Psychology , volume=
Taboo trade-offs: Reactions to transactions that transgress the spheres of justice , author=. Political Psychology , volume=. 1997 , publisher=
1997
-
[143]
Cognition , volume=
Reason-based choice , author=. Cognition , volume=. 1993 , publisher=
1993
-
[144]
Journal of Social Issues , year =
Deutsch, Morton , title =. Journal of Social Issues , year =
-
[145]
Journal of Economic Behavior & Organization , year =
Konow, James , title =. Journal of Economic Behavior & Organization , year =
-
[146]
Allan , title =
Lind, E. Allan , title =. Advances in Organizational Justice , editor =
-
[147]
and Weber, Roberto A
Krupka, Erin L. and Weber, Roberto A. , title =. Journal of the European Economic Association , year =
-
[148]
and Eavey, Cheryl L
Frohlich, Norman and Oppenheimer, Joe A. and Eavey, Cheryl L. , title =. British Journal of Political Science , year =
-
[149]
American Psychologist , author =
The construction of preference , volume =. American Psychologist , author =. 1995 , note =. doi:10.1037/0003-066X.50.5.364 , abstract =
1995 doi
-
[150]
Value in Health , author =
Multidimensional. Value in Health , author =. 2024 , keywords =. doi:10.1016/j.jval.2024.02.009 , abstract =
2024 doi
-
[151]
Social Science & Medicine , author =
Participatory health system priority setting:. Social Science & Medicine , author =. 2015 , keywords =. doi:10.1016/j.socscimed.2015.10.042 , abstract =
2015 doi
-
[152]
Participatory
Jain, Pallavi and Sornat, Krzysztof and Talmon, Nimrod , month = jul, year =. Participatory. doi:10.24963/ijcai.2020/54 , abstract =
2020 doi
-
[153]
Zeng, Xiong and Yu, Jing and Ozay, Necmiye , doi =. System. 2025 , bdsk-url-1 =
2025
-
[154]
Journal of Management Accounting Research , author =
Agency. Journal of Management Accounting Research , author =. 2009 , pages =. doi:10.2308/jmar.2009.21.1.317 , abstract =
2009 doi
- [155]
-
[156]
Journal of Mathematical Psychology , author =
Revealed. Journal of Mathematical Psychology , author =. 2002 , keywords =. doi:10.1006/jmps.2001.1398 , abstract =
2002
-
[157]
Eliciting
Camara, Modibo and Immorlica, Nicole and Lucier, Brendan , file =. Eliciting
-
[158]
Canadian Family Physician , author =
Eliciting patient values and preferences to inform shared decision making in preventive screening , volume =. Canadian Family Physician , author =. 2018 , pmid =
2018
-
[159]
Fair and
Wellings, Thomas and Banaie Heravan, Fatemeh and Sharma, Abhinav and Gelauff, Lodewijk and Fricker, Regula Hänggli and Pournaras, Evangelos , editor =. Fair and. Design for. 2024 , pages =. doi:10.1007/978-3-031-61698-3_6 , abstract =
2024 doi
-
[160]
Journal of Multi-Criteria Decision Analysis , author =
A. Journal of Multi-Criteria Decision Analysis , author =. 2015 , note =. doi:10.1002/mcda.1544 , abstract =
2015 doi
-
[161]
Distortion
Flanigan, Bailey and Procaccia, Ariel D and Wang, Sven , month = jul, year =. Distortion. Proceedings of the 24th. doi:10.1145/3580507.3597722 , abstract =
-
[162]
Journal of Risk and Uncertainty , author =
Determinants of stated willingness to pay for public goods:. Journal of Risk and Uncertainty , author =. 1994 , keywords =. doi:10.1007/BF01073401 , abstract =
1994 doi
- [163]
-
[164]
arXiv preprint arXiv:2406.11191 , year=
A survey on human preference learning for large language models , author=. arXiv preprint arXiv:2406.11191 , year=
-
[165]
Proceedings of the 2024 ACM Conference on Fairness, Accountability, and Transparency , pages=
Collective constitutional ai: Aligning a language model with public input , author=. Proceedings of the 2024 ACM Conference on Fairness, Accountability, and Transparency , pages=
2024
-
[166]
Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , volume =
On The Stability of Moral Preferences: A Problem with Computational Elicitation Methods , author =. Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , volume =. 2024 , doi =
2024
-
[167]
Advances in Neural Information Processing Systems , volume =
Yan, Songbai and Chaudhuri, Kamalika and Javidi, Tara , title =. Advances in Neural Information Processing Systems , volume =
-
[168]
arXiv preprint arXiv:2402.05070 , year=
A roadmap to pluralistic alignment , author=. arXiv preprint arXiv:2402.05070 , year=
-
[169]
Advances in neural information processing systems , volume=
Training language models to follow instructions with human feedback , author=. Advances in neural information processing systems , volume=
-
[170]
arXiv preprint arXiv:2303.16755 , year=
Training language models with language feedback at scale , author=. arXiv preprint arXiv:2303.16755 , year=
-
[171]
arXiv preprint arXiv:2311.02242 , year=
Democratic policy development using collective dialogues and AI , author=. arXiv preprint arXiv:2311.02242 , year=
-
[172]
Recerca: revista de pensament i an
Polis: Scaling deliberation by mapping high dimensional opinion spaces , author=. Recerca: revista de pensament i an. 2021 , publisher=
2021
-
[173]
Science , volume=
AI can help humans find common ground in democratic deliberation , author=. Science , volume=. 2024 , publisher=
2024
-
[174]
Journal of the ACM , volume=
Generative social choice , author=. Journal of the ACM , volume=. 2026 , publisher=
2026
-
[175]
Judgment and Decision Making , volume=
Decision conflict drives reaction times and utilitarian responses in sacrificial dilemmas , author=. Judgment and Decision Making , volume=. 2019 , publisher=
2019
-
[176]
Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , volume=
On the Pros and Cons of Active Learning for Moral Preference Elicitation , author=. Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , volume=
-
[177]
The Journal of Chemical Physics , volume =
Equation of State Calculations by Fast Computing Machines , author =. The Journal of Chemical Physics , volume =. 1953 , doi =
1953
-
[178]
Biometrika , volume =
Monte Carlo Sampling Methods Using Markov Chains and Their Applications , author =. Biometrika , volume =. 1970 , doi =
1970
-
[179]
Journal of the American Statistical Association , volume =
Sampling-Based Approaches to Calculating Marginal Densities , author =. Journal of the American Statistical Association , volume =. 1990 , doi =
1990
-
[180]
Handbook of Simulation Optimization , series =
A Guide to Sample Average Approximation , author =. Handbook of Simulation Optimization , series =. 2015 , doi =
2015
-
[181]
The Annals of Statistics , volume =
Markov Chains for Exploring Posterior Distributions , author =. The Annals of Statistics , volume =. 1994 , doi =
1994
-
[182]
The American Statistician , volume =
Understanding the Metropolis-Hastings Algorithm , author =. The American Statistician , volume =. 1995 , doi =
1995
-
[183]
Operations Research , volume =
Efficient Monte Carlo Procedures for Generating Points Uniformly Distributed over Bounded Regions , author =. Operations Research , volume =. 1984 , doi =
1984
-
[184]
Advances in Neural Information Processing Systems 31 , pages =
Maximizing Acquisition Functions for Bayesian Optimization , author =. Advances in Neural Information Processing Systems 31 , pages =. 2018 , publisher =
2018
-
[185]
arXiv preprint arXiv:2405.02612 , year=
Learning linear utility functions from pairwise comparison queries , author=. arXiv preprint arXiv:2405.02612 , year=
-
[186]
and Lovett, Shachar and Moran, Shay and Zhang, Jiapeng , title =
Kane, Daniel M. and Lovett, Shachar and Moran, Shay and Zhang, Jiapeng , title =. 2017 IEEE 58th Annual Symposium on Foundations of Computer Science (FOCS) , pages =. 2017 , doi =
2017
-
[187]
arXiv preprint arXiv:2511.10032 , year =
Moral Change or Noise? On Problems of Aligning AI With Temporally Unstable Human Feedback , author =. arXiv preprint arXiv:2511.10032 , year =. doi:10.48550/arXiv.2511.10032 , note =
-
[188]
Philosophical Studies , year =
Beyond Preferences in AI Alignment , author =. Philosophical Studies , year =. doi:10.1007/s11098-024-02249-w , url =
-
[190]
arXiv preprint arXiv:2505.22820 , year =
Preference Learning with Response Time , author =. arXiv preprint arXiv:2505.22820 , year =
-
[191]
Davidson , title =
Roger R. Davidson , title =. Biometrics , year =
-
[192]
H. A. David , title =. 1988 , series =
1988
-
[193]
Davidson and Peter H
Roger R. Davidson and Peter H. Farquhar , title =. Biometrics , year =
-
[194]
Biometrika , year =
David Firth , title =. Biometrika , year =
-
[195]
Journal of Statistical Software , year =
David Firth , title =. Journal of Statistical Software , year =
-
[196]
2008 , organization =
David Firth , title =. 2008 , organization =
2008
-
[197]
Advances in neural information processing systems , volume=
Occam's razor is insufficient to infer the preferences of irrational agents , author=. Advances in neural information processing systems , volume=
-
[198]
Psychometrika , year =
Frederick Mosteller , title =. Psychometrika , year =
-
[199]
Journal of Artificial Intelligence Research , volume =
Value Preferences Estimation and Disambiguation in Hybrid Participatory Systems , author =. Journal of Artificial Intelligence Research , volume =. 2025 , doi =
2025
-
[200]
HHAI 2022: Augmenting Human Intellect , series =
Estimating Value Preferences in a Hybrid Participatory System , author =. HHAI 2022: Augmenting Human Intellect , series =. 2022 , doi =
2022
-
[201]
Advances in Neural Information Processing Systems , volume =
Enhancing Preference-based Linear Bandits via Human Response Time , author =. Advances in Neural Information Processing Systems , volume =. 2024 , doi =
2024
-
[202]
13th Multidisciplinary Workshop on Adv
Learning Individual and Collective Priorities over Moral Dilemmas with the Life Jacket Dataset , author=. 13th Multidisciplinary Workshop on Adv. in Preference Handling, Vienna, Austria , year=
-
[203]
AI Magazine , author =
Preference. AI Magazine , author =. 2008 , note =. doi:10.1609/aimag.v29i4.2201 , abstract =
2008 doi
-
[204]
Journal of Theoretical Politics , author =
A. Journal of Theoretical Politics , author =. 2000 , keywords =
2000
-
[205]
AI Magazine , author =
Elicitation of. AI Magazine , author =. 2008 , note =. doi:10.1609/aimag.v29i4.2203 , abstract =
2008 doi
- [206]
-
[207]
Proceedings of the 2025 CHI Conference on Human Factors in Computing Systems , pages=
Can ai model the complexities of human moral decision-making? a qualitative study of kidney allocation decisions , author=. Proceedings of the 2025 CHI Conference on Human Factors in Computing Systems , pages=
2025
-
[208]
and Vossler, Patrick and Blessenohl, Simon and Vayanos, Phebe , month = oct, year =
Johnston, Caroline M. and Vossler, Patrick and Blessenohl, Simon and Vayanos, Phebe , month = oct, year =. Deploying a. Equity and. doi:10.1145/3617694.3623254 , abstract =
-
[209]
Trading mental effort for confidence in the metacognitive control of value-based decision-making
-
[211]
British Medical Bulletin , author =
Ordinal preference elicitation methods in health economics and health services research: using discrete choice experiments and ranking methods , volume =. British Medical Bulletin , author =. 2012 , pmid =. doi:10.1093/bmb/lds020 , abstract =
2012 doi
-
[212]
Preference
Hsee, Christopher K and Loewenstein, George F and Blount, Sally and Bazerman, Max H , file =. Preference
-
[213]
and Holyoak, Keith J
Simon, Dan and Krawczyk, Daniel C. and Holyoak, Keith J. , month = jul, year =. Construction of
-
[214]
ResearchGate , doi =
(. ResearchGate , doi =
-
[215]
arXiv.org , author =
On the. arXiv.org , author =. 2024 , file =
2024
-
[216]
and Brusse, Ciell and Velgersdijk, Ronald and Zhu, Haiyi , month = jun, year =
Shen, Hong and Wang, Leijie and Deng, Wesley H. and Brusse, Ciell and Velgersdijk, Ronald and Zhu, Haiyi , month = jun, year =. The. Proceedings of the 2022. doi:10.1145/3531146.3533110 , abstract =
2022
-
[217]
Qual Health Care , author =
Use of discrete choice experiments to elicit preferences , abstract =. Qual Health Care , author =. 2025 , file =
2025
-
[218]
Transportation Research Part F: Traffic Psychology and Behaviour , author =
Decision modeling for automated driving in dilemmas based on bidirectional value alignment of moral theory values and fair human moral values , volume =. Transportation Research Part F: Traffic Psychology and Behaviour , author =. 2025 , keywords =. doi:10.1016/j.trf.2024.11.0...
2025 doi
-
[219]
Journal of Political Economy , author =
A. Journal of Political Economy , author =. 2020 , pages =. doi:10.1086/706861 , language =
2020 doi
-
[220]
The Quarterly Journal of Economics , author =
A. The Quarterly Journal of Economics , author =. 2014 , pages =. doi:10.1093/qje/qju024 , abstract =
2014 doi
-
[221]
Advances in psychology , volume=
Decision rules and the search for a dominance structure: Towards a process model of decision making , author=. Advances in psychology , volume=. 1983 , publisher=
1983
-
[222]
arXiv preprint arXiv:2505.23749 , year=
Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? , author=. arXiv preprint arXiv:2505.23749 , year=
-
[223]
Journal of Experimental Psychology , year =
Reversals of Preference Between Bids and Choices in Gambling Decisions , author =. Journal of Experimental Psychology , year =
-
[224]
The Quarterly Journal of Economics , year =
Beyond Revealed Preference: Choice-Theoretic Foundations for Behavioral Welfare Economics , author =. The Quarterly Journal of Economics , year =
-
[225]
Perspectives on Psychological Science , year =
The Inversion Problem: Why Algorithms Should Infer Mental State and Not Just Predict Behavior , author =. Perspectives on Psychological Science , year =
-
[226]
Journal of Political Economy , year =
An Economic Theory of Self-Control , author =. Journal of Political Economy , year =
-
[227]
Econometrica , year =
Temptation and Self-Control , author =. Econometrica , year =
-
[228]
arXiv preprint arXiv:1309.6864 , year=
Preference elicitation for general random utility models , author=. arXiv preprint arXiv:1309.6864 , year=
-
[229]
Advances in Neural Information Processing Systems , year =
Deep Reinforcement Learning from Human Preferences , author =. Advances in Neural Information Processing Systems , year =
-
[230]
Advances in Neural Information Processing Systems , year =
Cooperative Inverse Reinforcement Learning , author =. Advances in Neural Information Processing Systems , year =
-
[231]
2020 , journal =
What If I Don't Like Any Of The Choices? The Limits of Preference Elicitation for Participatory Algorithm Design , author =. 2020 , journal =
2020
-
[232]
Journal of Applied Psychology , volume =
On the Dimensionality of Organizational Justice: A Construct Validation of a Measure , author =. Journal of Applied Psychology , volume =. 2001 , doi =
2001
-
[233]
American Economic Review , volume=
Individual preferences for giving , author=. American Economic Review , volume=. 2007 , publisher=
2007
-
[234]
Yue, Yisong and Broder, Josef and Kleinberg, Robert and Joachims, Thorsten , journal =. The. 2012 , volume =
2012
-
[235]
Advances in Neural Information Processing Systems , year =
Copeland Dueling Bandits , author =. Advances in Neural Information Processing Systems , year =
-
[236]
Journal of Machine Learning Research , year =
Preference-based Online Learning with Dueling Bandits: A Survey , author =. Journal of Machine Learning Research , year =
-
[237]
, journal =
Lee, Min Kyung and Kusbit, Daniel and Kahng, Anson and Kim, Ji Tae and Yuan, Xinran and Chan, Allissa and See, Daniel and Noothigattu, Ritesh and Lee, Siheon and Psomas, Alexandros and Procaccia, Ariel D. , journal =. 2019 , publisher =
2019
-
[238]
American Psychologist , year =
Slovic, Paul , title =. American Psychologist , year =
-
[239]
Science , year =
Tversky, Amos and Kahneman, Daniel , title =. Science , year =
-
[240]
American Psychologist , year =
Kahneman, Daniel and Tversky, Amos , title =. American Psychologist , year =
-
[241]
and Bettman, James R
Payne, John W. and Bettman, James R. and Johnson, Eric J. , title =
-
[242]
Mathematics of operations research , volume=
Minimization by random search techniques , author=. Mathematics of operations research , volume=. 1981 , publisher=
1981
-
[243]
Management Science , year =
Tversky, Amos and Simonson, Itamar , title =. Management Science , year =
-
[244]
Extending the Bounds of Rationality: Evidence and Theories of Preferential Choice , journal =
Rieskamp, J. Extending the Bounds of Rationality: Evidence and Theories of Preferential Choice , journal =. 2006 , volume =
2006
-
[245]
Festinger, Leon , title =
-
[246]
and Berntson, Gary G
Cacioppo, John T. and Berntson, Gary G. , title =. Psychological Bulletin , year =
-
[247]
and Petty, Richard E
Priester, Joseph R. and Petty, Richard E. , title =. Journal of Personality and Social Psychology , year =
-
[248]
, title =
van Harreveld, Frenk and van der Pligt, Joop and de Liver, Yael N. , title =. Personality and Social Psychology Review , year =
-
[249]
and Schwarz, Norbert , title =
Schneider, Iris K. and Schwarz, Norbert , title =. Current Opinion in Behavioral Sciences , year =
-
[250]
Luttrell, Andrew and Petty, Richard E. and Bri. Ambivalence and Certainty Can Interact in Predicting Attitude Stability Over Time , journal =. 2016 , doi =
2016
-
[251]
and Braver, Todd S
Botvinick, Matthew M. and Braver, Todd S. and Barch, Deanna M. and Carter, Cameron S. and Cohen, Jonathan D. , title =. Psychological Review , year =
-
[252]
and Cohen, Jonathan D
Shenhav, Amitai and Botvinick, Matthew M. and Cohen, Jonathan D. , title =. Neuron , year =
-
[253]
, title =
Anderson, Christopher J. , title =. Psychological Bulletin , year =
-
[254]
Journal of Consumer Research , year =
Dhar, Ravi , title =. Journal of Consumer Research , year =
-
[255]
Psychological Science , year =
Tversky, Amos and Shafir, Eldar , title =. Psychological Science , year =
-
[256]
and Mann, Leon , title =
Janis, Irving L. and Mann, Leon , title =
-
[257]
, title =
Hogarth, Robin M. , title =
-
[258]
and Townsend, James T
Busemeyer, Jerome R. and Townsend, James T. , title =. Psychological Review , year =
-
[259]
and Brown, Scott D
Ratcliff, Roger and Smith, Philip L. and Brown, Scott D. and McKoon, Gail , title =. Trends in Cognitive Sciences , year =
-
[260]
and van der Vaart, Aad W
Ghosal, Subhashis and Ghosh, Jayanta K. and van der Vaart, Aad W. , title =. The Annals of Statistics , volume =
-
[261]
, title =
Walker, Stephen G. , title =. The Annals of Statistics , volume =
-
[262]
Zeitschrift f
Schwartz, Lorraine , title =. Zeitschrift f
-
[263]
Advances in Neural Information Processing Systems , volume=
Deep reinforcement learning from human preferences , author=. Advances in Neural Information Processing Systems , volume=
-
[264]
arXiv preprint arXiv:1909.08593 , year=
Fine-tuning language models from human preferences , author=. arXiv preprint arXiv:1909.08593 , year=
1909 arXiv
-
[265]
Advances in Neural Information Processing Systems , volume=
Learning to summarize from human feedback , author=. Advances in Neural Information Processing Systems , volume=
-
[266]
Advances in Neural Information Processing Systems , volume=
Training language models to follow instructions with human feedback , author=. Advances in Neural Information Processing Systems , volume=
-
[267]
arXiv preprint arXiv:2305.18290 , year=
Direct preference optimization: Your language model is secretly a reward model , author=. arXiv preprint arXiv:2305.18290 , year=
-
[268]
arXiv preprint , year=
Understanding moral preferences in kidney allocation: A qualitative study , author=. arXiv preprint , year=
-
[269]
2024 , eprint =
Reward Learning From Preference With Ties , author =. 2024 , eprint =
2024
-
[270]
Proceedings of EMNLP , year=
Ties matter: Meta-evaluating modern metrics and human preferences in machine translation , author=. Proceedings of EMNLP , year=
-
[271]
arXiv preprint arXiv:2402.17257 , year=
RIME: Robust preference-based reinforcement learning with noisy preferences , author=. arXiv preprint arXiv:2402.17257 , year=
-
[272]
Findings of EMNLP , year=
Characterizing and mitigating ranking errors in human preference data , author=. Findings of EMNLP , year=
-
[273]
Proceedings of the 35th International Conference on Machine Learning , pages=
Online learning with abstention , author=. Proceedings of the 35th International Conference on Machine Learning , pages=
-
[274]
arXiv preprint arXiv:2402.14585 , year=
Bandits with abstention under expert advice , author=. arXiv preprint arXiv:2402.14585 , year=
-
[275]
International Conference on Learning Representations , year=
Reward model ensembles help mitigate overoptimization , author=. International Conference on Learning Representations , year=
-
[276]
arXiv preprint arXiv:2401.00243 , year=
Uncertainty-penalized reinforcement learning from human feedback with diverse reward ensembles , author=. arXiv preprint arXiv:2401.00243 , year=
-
[277]
arXiv preprint arXiv:2404.10271 , year=
Social choice should guide AI alignment in dealing with diverse human feedback , author=. arXiv preprint arXiv:2404.10271 , year=
-
[278]
Advances in Neural Information Processing Systems , year=
Axioms for AI alignment from human feedback , author=. Advances in Neural Information Processing Systems , year=
-
[279]
Proceedings of AAAI , year=
A voting-based system for ethical decision making , author=. Proceedings of AAAI , year=
-
[280]
Proceedings of the 38th International Conference on Machine Learning , year=
Value alignment verification , author=. Proceedings of the 38th International Conference on Machine Learning , year=
-
[281]
and Martin, Ryan and Ghosh, Jayanta K
Tokdar, Surya T. and Martin, Ryan and Ghosh, Jayanta K. , title =. The Annals of Statistics , volume =
-
[282]
arXiv preprint arXiv:1112.5745 , year=
Bayesian active learning for classification and preference learning , author=. arXiv preprint arXiv:1112.5745 , year=
-
[283]
Proceedings of the 22nd international conference on Machine learning , pages=
Preference learning with Gaussian processes , author=. Proceedings of the 22nd international conference on Machine learning , pages=
-
[284]
2017 , journal =
A Voting-Based System for Ethical Decision Making , author =. 2017 , journal =
2017
-
[285]
Active preference-based learning of reward functions , author=
-
[286]
Plos one , volume=
How should ICU beds be allocated during a crisis? Evidence from the COVID-19 pandemic , author=. Plos one , volume=. 2022 , publisher=
2022
-
[287]
arXiv preprint arXiv:2110.07574 , year=
Can machines learn morality? the delphi experiment , author=. arXiv preprint arXiv:2110.07574 , year=
-
[288]
Proceedings of the 2017 CHI conference on human factors in computing systems , pages=
A human-centered approach to algorithmic services: Considerations for fair and motivating smart community service management that allocates donations to non-profit organizations , author=. Proceedings of the 2017 CHI conference on human factors in computing systems , pages=
2017
-
[289]
arXiv preprint arXiv:2509.04445 , year=
Towards cognitively-faithful decision-making models to improve ai alignment , author=. arXiv preprint arXiv:2509.04445 , year=
-
[290]
value-action
What shapes the “value-action” gap?. Economia Politica , author =. 2022 , keywords =. doi:10.1007/s40888-022-00282-8 , abstract =
2022 doi
-
[291]
and Liscio, Enrico and Murukannaiah, Pradeep K
Siebert, Luciano C. and Liscio, Enrico and Murukannaiah, Pradeep K. and Kaptein, Lionel and Spruit, Shannon and Van Den Hoven, Jeroen and Jonker, Catholijn , editor =. Estimating. Frontiers in. 2022 , doi =
2022
-
[292]
Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Indecision modeling , author=. Proceedings of the AAAI Conference on Artificial Intelligence , volume=
-
[293]
, author=
A subjective utilitarian theory of moral judgment. , author=. Journal of Experimental Psychology: General , volume=. 2016 , publisher=
2016
-
[294]
Ethics , volume=
The additive fallacy , author=. Ethics , volume=. 1988 , publisher=
1988
-
[295]
Journal of Business Ethics , volume=
A comparison of models describing the impact of moral decision making on investment decisions , author=. Journal of Business Ethics , volume=. 2008 , publisher=
2008
-
[296]
Handbook of theories of social psychology , volume=
Construal level theory , author=. Handbook of theories of social psychology , volume=. 2012 , publisher=
2012
-
[297]
value-action
What shapes the “value-action” gap? The role of time perception reconsidered , author=. Economia Politica , volume=. 2022 , publisher=
2022
-
[298]
Proceedings of the AAAI Conference on Artificial Intelligence , volume=
Proportional aggregation of preferences for sequential decision making , author=. Proceedings of the AAAI Conference on Artificial Intelligence , volume=
-
[299]
, author=
The gradual threshold model of ambivalence: relating the positive and negative bases of attitudes to subjective ambivalence. , author=. Journal of personality and social psychology , volume=. 1996 , publisher=
1996
-
[300]
Perspectives on Psychological Science , volume=
Toward a general framework of biased reasoning: Coherence-based reasoning , author=. Perspectives on Psychological Science , volume=. 2025 , publisher=
2025
Reviewed August 2, 2026 · model on record in the stance chip above.
Discussion (0). Sign in to comment.