Beyond Partisan Leaning: A Comparative Analysis of Political Bias in Large Language Models
arXiv:2412.16746 · doi:10.1080/19331681.2026.2646990
Abstract
As large language models (LLMs) become increasingly embedded in civic, educational, and political information environments, concerns about their potential political bias have grown. Prior research often evaluates such bias through simulated personas or predefined ideological typologies, which may introduce artificial framing effects or overlook how models behave in general use scenarios. This study adopts a persona-free, topic-specific approach to evaluate political behavior in LLMs, reflecting how users typically interact with these systems-without ideological role-play or conditioning. We introduce a two-dimensional framework: one axis captures partisan orientation on highly polarized topics (e.g., abortion, immigration), and the other assesses sociopolitical engagement on less polarized issues (e.g., climate change, foreign policy). Using survey-style prompts drawn from the ANES and Pew Research Center, we analyze responses from 43 LLMs developed in the U.S., Europe, China, and the Middle East. We propose an entropy-weighted bias score to quantify both the direction and consistency of partisan alignment, and identify four behavioral clusters through engagement profiles. Findings show most models lean center-left or left ideologically and vary in their nonpartisan engagement patterns. Model scale and openness are not strong predictors of behavior, suggesting that alignment strategy and institutional context play a more decisive role in shaping political expression.
References in corpus (16)
- On the Opportunities and Risks of Foundation Models
- Cultural Bias and Cultural Alignment of Large Language Models
- Towards Understanding Sycophancy in Language Models
- Holistic Evaluation of Language Models
- Whose Opinions Do Language Models Reflect?
- Large Language Models Can Be Used to Estimate the Latent Positions of Politicians
- Foundational Challenges in Assuring Alignment and Safety of Large Language Models
- CultureLLM: Incorporating Cultural Differences into Large Language Models
- From Pretraining Data to Language Models to Downstream Tasks: Tracking the Trails of Political Biases Leading to Unfair NLP Models
- OR-Bench: An Over-Refusal Benchmark for Large Language Models
- Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
- The Political Preferences of LLMs
- Mapping and Influencing the Political Ideology of Large Language Models using Synthetic Personas
- Measurement in the Age of LLMs: An Application to Ideological Scaling
- PoliTune: Analyzing the Impact of Data Selection and Fine-Tuning on Economic and Political Biases in Large Language Models
- Multilingual Political Views of Large Language Models: Identification and Steering