Pith. sign in

REVIEW 4 cited by

The Political Preferences of LLMs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.01789 v2 pith:UZLVG4N6 submitted 2024-02-02 cs.CY cs.AIcs.CL

classification cs.CYcs.AIcs.CL
keywords llmspoliticalpreferencesmodelsbaseconversationalembeddedorientation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

I report here a comprehensive analysis about the political preferences embedded in Large Language Models (LLMs). Namely, I administer 11 political orientation tests, designed to identify the political preferences of the test taker, to 24 state-of-the-art conversational LLMs, both closed and open source. When probed with questions/statements with political connotations, most conversational LLMs tend to generate responses that are diagnosed by most political test instruments as manifesting preferences for left-of-center viewpoints. This does not appear to be the case for five additional base (i.e. foundation) models upon which LLMs optimized for conversation with humans are built. However, the weak performance of the base models at coherently answering the tests' questions makes this subset of results inconclusive. Finally, I demonstrate that LLMs can be steered towards specific locations in the political spectrum through Supervised Fine-Tuning (SFT) with only modest amounts of politically aligned data, suggesting SFT's potential to embed political orientation in LLMs. With LLMs beginning to partially displace traditional information sources like search engines and Wikipedia, the societal implications of political biases embedded in LLMs are substantial.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 6 citations worldwide. Full citation record

  1. BiasLab: A Multilingual Dual-Framing Framework for LLM Bias Measurement, Applied to Workplace and HR Contexts

    cs.CL 2026-01 reject novelty 5.0 of 10

    BiasLab uses mirrored affirmative/reverse prompt pairs across 12 languages to quantify directional preferences in 10 LLMs, claiming a systematic asymmetry between rejection and endorsement.

  2. Normative Evaluation of Large Language Models with Everyday Moral Dilemmas

    cs.AI 2025-01 conditional novelty 5.0 of 10

    Seven LLMs give different moral verdicts on AITA dilemmas, differ from Redditors, and only in an ensemble approximate human consensus.

  3. Rethinking stance detection: A theoretically-informed research agenda for user-level inference using language models

    cs.CL 2025-02 accept novelty 4.0 of 10

    A theoretically grounded framework and agenda for shifting stance detection from message-level labels to user-level inference using LLM-inferred psychological attributes.

  4. A Survey of Large Language Models in Discipline-specific Research: Challenges, Methods and Opportunities

    cs.CL 2025-07 conditional novelty 2.0 of 10

    A review that categorizes methods for adapting LLMs to discipline-specific research and surveys applications across five broad academic fields.

Pith tools