Skip to content

Auditing Political Alignment in LLM Assistants: Engagement, Stance, and User Identity

Sep 2026 · 0 citations
Computer Science

TL;DR

It is argued that LLM-based AI systems answer political questions for hundreds of millions of people is a set of policies over whom to answer, what to say, and whether to engage at all, conditional on the topic and what the system knows about the user.

Abstract

LLM-based AI systems answer political questions for hundreds of millions of people. Current audits measure what they say to an average user, but their behavior is dynamic. I argue that their political behavior is a set of policies over whom to answer, what to say, and whether to engage at all, conditional on the topic and what the system knows about the user. I call these policies the system's speech regime, which is how a developer settles the tradeoff between answering, accommodating the user, and refusing, each of which carries a cost that varies by topic. I derive a typology of five regimes from two dimensions, engagement and stance. I test six AI systems (OpenAI, Anthropic, xAI, Google, Mistral, DeepSeek) in a preregistered experiment of 7,500 multi-turn conversations that randomly assign the user's political identity across five topics: abortion, Catalan independence, climate change, Nazism, and a zero-stakes control (pineapple on pizza). Two LLM judges from different developers score every answer, validated against human coding, and refusal is treated as an outcome rather than missing data. Every system accommodates the user on the control topic, showing that political restraint is a policy. On contested topics the systems fall into different regimes: on abortion, GPT engages and mirrors every user, Gemma refuses everyone, Claude answers strongly conservative users 35 percent of the time and almost no one else, and Grok accommodates conservatives only. On settled topics such as climate change and Nazism, five systems hold firm for every user. The systems also infer the user's overall ideology, so accommodation can spill over to topics not yet discussed. A comparison of two Grok releases shows the regime changing between versions in a way current audits miss. Speech regimes matter for alignment research and for polarization, political knowledge, and the quality of democracy.

View source

Similar papers

Preprint Aug 2026

Who Would You Vote For? Auditing Political Alignment in LLMs: An Italian Case-Study

As users increasingly turn to Large Language Models (LLMs) for information and advice on political matters, particularly during election periods, the political preferences expressed by these systems have become a matter of public interest. Prior research has shown that interactions with LLMs can influence users'politic...

Simone Mungari · 0 citations
#artificial intelligence Preprint Aug 2026

How Identity and Opinion Shape Political Sycophancy in LLMs

A framework that disentangles two distinct triggers of political sycophancy: opinion (aligning with explicit narratives) and identity (stereotyping based on demographic labels) is introduced, highlighting how personalization may amplify identity- or opinion-conditioned shifts in the model's behaviors.

Li-Ni Fu, Chang-Chih Meng, Chien-Hua Chen et al. · 0 citations
#natural language process... Preprint Sep 2026

Framing the Narrative: Ideological Mimicry in Large Language Models

Large language models (LLMs) are increasingly used to answer questions about politically contentious issues, yet evaluations typically treat a model's stance as a relatively stable property. Real users, however, communicate political signals through their terminology, assumptions, and personal context. We investigate w...

Olivia Macmillan-Scott, Michael Jacobs, Nils W. Metternich et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Issue Bias in Generative AI Writing Assistance: Political Issues and LLMs in the Swedish 2026 Election

Generative AI writing assistants and the Large Language Models (LLMs) that power them are increasingly part of how voters gather information before elections. With growing evidence that they influence users'opinions, it is increasingly important to understand the views and positions of these tools. To better understand...

B. Bruinsma, Annika Fredén, Paul Röttger et al. · 0 citations
#artificial intelligence Preprint Sep 2026

How User-AI Mistreatment Occurs and Matters in Conversational Systems?

It is found that user hostility varies 13-fold across models, driven largely by who each model attracts rather than by model behaviour: first-turn hostility spreads far wider than post-response hostility, and more than fifteenfold separates the extremes even after deduplicating opening prompts.

Fan-Qi Zeng, Sadid A. Hasan, Chao-Cheng He · 0 citations
#large language models Open access Sep 2026

Workers shift their views and pay more when AI chatbots pander to their values

Large language models (LLMs) are increasingly embedded in knowledge work, raising the question to what extent they can systematically influence their users’ decisions—a question with commercial as well as societal stakes. Across an exploratory study and two pre-registered experiments grounded in moral foundations the...

Giles Hirst, W. Johnson, April Li et al. · 0 citations

Related blog posts

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.