Origin of Political Bias in Large Language Models
Determine the origins of the left-leaning political bias observed in large language models when evaluated on the Wahl-O-Mat political statements by identifying and quantifying the contributions of potential sources such as training data bias, representation gaps, model memory effects, and tokenizer-induced skew.
References
It is not certain where the bias originates, but a reasonable estimate would be an inherent bias in the training data.
— Large Means Left: Political Bias in Large Language Models Increases with Their Number of Parameters
(2505.04393 - Exler et al., 7 May 2025) in Discussion, paragraph 3
Tokenisation and final-layer geometry may contribute as well, but we controlled for neither and leave the architecture gap open.
— Reference-Based Bias Detection in LLMs via Relative Representations of Hidden States
(2609.10060 - Jeliński et al., 9 Sep 2026) in Section 4.1, “Fine-Tuning-Induced Representational Shifts on WildGuardMix”