arXiv · 2601.17016
Measuring Political Stance and Consistency in Large Language Models
Abstract
With the incredible advancements in Large Language Models (LLMs), many people have started using them to satisfy their information needs. However, utilizing LLMs might be problematic for political issues where disagreement is common and model outputs may reflect training-data biases or deliberate alignment choices. To better characterize such behavior, we assess the stances of nine LLMs on 24 politically sensitive issues using five prompting techniques. We find that models often adopt opposing stances on several issues; some positions are malleable under prompting, while others remain stable. Among the models examined, Grok-3-mini is the most persistent, whereas Mistral-7B is the least. For issues involving countries with different languages, models tend to support the side whose language is used in the prompt. Notably, no prompting technique alters model stances on the Qatar blockade or the oppression of Palestinians. We hope these findings raise user awareness when seeking political guidance from LLMs and encourage developers to address these concerns.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Salah Feras Alali, Mohammad Nashat Maasfeh, Mucahid Kutlu, Saban Kardas. 2026-01-15. Measuring Political Stance and Consistency in Large Language Models. https://arxiv.org/abs/2601.17016
Cite the original work for its findings. Save a collection to share your selection of sources.