REVIEW 9 cited by
LLM Voting: Human Choices and AI Collective Decision Making
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This paper investigates the voting behaviors of Large Language Models (LLMs), specifically GPT-4 and LLaMA-2, their biases, and how they align with human voting patterns. Our methodology involved using a dataset from a human voting experiment to establish a baseline for human preferences and conducting a corresponding experiment with LLM agents. We observed that the choice of voting methods and the presentation order influenced LLM voting outcomes. We found that varying the persona can reduce some of these biases and enhance alignment with human choices. While the Chain-of-Thought approach did not improve prediction accuracy, it has potential for AI explainability in the voting process. We also identified a trade-off between preference diversity and alignment accuracy in LLMs, influenced by different temperature settings. Our findings indicate that LLMs may lead to less diverse collective outcomes and biased assumptions when used in voting scenarios, emphasizing the need for cautious integration of LLMs into democratic processes.
Forward citations
Cited by 9 Pith papers
-
Test-Time Scaling via Error Localization
TTEL uses feedback-induced token probability drops to localize the first error in a failed reasoning trace and branch a new generation from that prefix, improving pass@k per token on coding and math benchmarks.
-
LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra
The LLM Economist framework couples persona-conditioned worker agents with an in-context RL planner to search US-bracket tax schedules, yet its Saez benchmark is derived from the planner's own solution and its headlin...
-
New Evaluation Paradigm for Lexical Simplification
This paper proposes an all-in-one lexical simplification dataset with per-sentence complex word and substitute annotations, and a multi-LLM voting method that is claimed to outperform earlier baselines.
-
Political Actor Agent: Simulating Legislative System for Roll Call Votes Prediction with Large Language Models
PAA, a role-playing LLM agent with multi-view planning and leader-follower influence, reports 91.8-92.1% accuracy on U.S. House roll-call prediction.
-
Decision Protocols in Multi-Agent Large Language Model Conversations
Consensus decision protocols beat voting/judge on knowledge QA for Llama-3 multi-agent chats, while voting and judge win on logic tasks; independent initial drafts raise accuracy and extra voting-time info barely helps.
-
A theory of appropriateness with applications to generative artificial intelligence
A theory that human and AI behavior is guided by context-dependent appropriateness implemented as predictive pattern completion, with norms as conventional sanctioning patterns.
-
Anger Speaks Louder? Exploring the Effects of AI Nonverbal Emotional Cues on Human Decision Certainty in Moral Dilemmas
Anger-themed animated chat balloons from an AI assistant increased reversal of decision certainty in moral dilemmas, whereas sadness cues did not.
-
Political-LLM: Large Language Models in Political Science
A survey and taxonomy of LLM applications in political science, with a case study suggesting that larger LLMs reproduce ANES 2016 voting patterns more accurately than smaller ones.
-
Towards unearthing neglected climate innovations from scientific literature using Large Language Models
A GPT-4o-based screening workflow with contextual prompting and control-fitted weighting can rank known climate spin-out abstracts highly, though validation is limited to a small sample.
Discussion (0). Continue with ORCID to comment.