Aligning Large Language Models with Diverse Political Viewpoints (2024.emnlp-main)
Copied to clipboard
| Challenge: | Large language models such as ChatGPT exhibit striking political biases . a recent study shows that chatbots exhibit progressive, liberal, and proenvironmental biase . |
| Approach: | They propose to align large language models with 100,000 comments from candidates running for national parliament in Switzerland. |
| Outcome: | The proposed model generates more accurate political viewpoints from Swiss parties compared to commercial models such as ChatGPT. |
Similar Papers
Aligning Language Models to User Opinions (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Personality is a defining feature of human beings, shaped by a complex interplay of demographic characteristics, moral principles, and social experiences. |
| Approach: | They use public opinion surveys to model past user opinions in addition to user demographics and ideology to achieve up to 7 points accuracy gains in predicting public opinions from survey questions. |
| Outcome: | The proposed model achieves 7 points accuracy gains in predicting public opinions from public opinion surveys across a broad set of topics. |
Navigating the Political Compass: Evaluating Multilingual LLMs across Languages and Nationalities (2025.findings-acl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) are ubiquitous in today’s technological landscape, boasting a plethora of applications, and even endangering human jobs in complex and creative fields. |
| Approach: | They evaluate the political bias of 15 multilingual LLMs using the Political Compass Test and assign a nationality to each model. |
| Outcome: | The models on the 50 most populous countries and their official languages exhibit political bias. |
Quite Good, but Not Enough: Nationality Bias in Large Language Models - a Case Study of ChatGPT (2024.lrec-main)
Copied to clipboard
| Challenge: | Nationality is a key demographic element that enhances the performance of large language models, but it has received less scrutiny regarding inherent biases. |
| Approach: | They investigated nationality bias in ChatGPT, a large language model for text generation. |
| Outcome: | The proposed model generates 4,680 discourses about nationality in Chinese and English, with 195 countries, 4 temperature settings, and 3 prompt types. |
Analyzing Political Bias in LLMs via Target-Oriented Sentiment Classification (2025.findings-acl)
Copied to clipboard
| Challenge: | Existing methods to analyze political biases rely on small-size intermediate tasks and the LLMs themselves. |
| Approach: | They propose an entropy-based inconsistency metric to encode political biases . they insert 1319 demographically and politically diverse politician names in 450 political sentences . |
| Outcome: | The proposed method combines high accuracy with a correct understanding of the candidate candidate. |
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said (2024.acl-long)
Copied to clipboard
| Challenge: | Existing benchmarks and measures focus on gender and racial biases, but political bias exists in LLMs and can lead to polarization and other harms in downstream applications. |
| Approach: | They propose to analyze the content and style of LLMs generated by political issues and propose a framework that can be scalable to other topics. |
| Outcome: | The proposed framework is easily scalable to other topics and is explainable. |
LLM Tropes: Revealing Fine-Grained Values and Opinions in Large Language Models (2024.findings-emnlp)
Copied to clipboard
| Challenge: | Existing approaches to evaluate latent values and opinions in large language models suffer from three notable shortcomings. |
| Approach: | They propose to analyze 156k LLM responses to 62 propositions of the Political Compass Test (PCT) generated by 6 LLMs using 420 prompt variations. |
| Outcome: | The proposed analysis of 156k LLM responses to the Political Compass Test (PCT) generated by 6 LLMs shows that tropes are recurrent and consistent across prompts. |
Exploiting contextual information to improve stance detection in informal political discourse with LLMs (2025.acl-srw)
Copied to clipboard
| Challenge: | Political stance detection is an increasingly relevant part of analyzing the flow of ideas in online environments where discourse is informal and implicitly expressed. |
| Approach: | They evaluate large language models for political stance detection in informal online discourse by analyzing user profiles derived from historical posts. |
| Outcome: | The proposed model improves accuracy by up to 74% on a political forum dataset. |
Llama meets EU: Investigating the European political spectrum through the lens of LLMs (2024.naacl-short)
Copied to clipboard
| Challenge: | Large Language Models inherit clear political leanings that have been shown to influence downstream task performance. |
| Approach: | They adapt Llama Chat to a European political context and audit its political leanings based on the EUandI questionnaire to analyze its political knowledge and ability to reason in context. |
| Outcome: | The proposed model is adapted from speeches of individual euro-parties from debates in the European Parliament to analyze its political leanings. |
Quantifying Generative Media Bias with a Corpus of Real-world and Generated News Articles (2024.findings-emnlp)
Copied to clipboard
| Challenge: | Existing studies focus on LLMs undertaking political questionnaires, which offers only limited insights into their biases and operational nuances. |
| Approach: | They propose to use a curated dataset to generate 56,700 synthetic articles using nine LLMs. |
| Outcome: | The proposed model can detect political biases using supervised models and LLMs. |
Algorithmic Fidelity of Large Language Models in Generating Synthetic German Public Opinions: A Case Study (2025.acl-long)
Copied to clipboard
Bolei Ma, Berk Yoztyurk, Anna-Carolina Haensch, Xinpeng Wang, Markus Herklotz, Frauke Kreuter, Barbara Plank, Matthias Aßenmacher
| Challenge: | Recent advances in large language models have generated significant interest in their potential for synthetic data generation across various domains. |
| Approach: | They use open-ended survey data from the German Longitudinal Election Studies to prompt different LLMs to generate synthetic public opinions reflective of German subpopulations by incorporating demographic features into the persona prompts. |
| Outcome: | The LLM performs better for supporters of left-leaning parties like The Greens and The Left compared to other parties, and matches the least with the right-party AfD. |