Evaluating Credibility and Political Bias in LLMs for News Outlets in Bangladesh (2025.acl-srw)
Copied to clipboard
| Challenge: | Large language models (LLMs) are widely used in search engines to provide direct an-swers, while AI chatbots retrieve updated infor-mation from the web. |
| Approach: | They audit nine Large Language Models from OpenAI, Google, and Meta to assess their ability to eval-uate the credibility and political bias of the top20 most popular news outlets in Bangladesh. |
| Outcome: | The proposed models show internal consistency in credibil-ity ratings, but misalignment with human experts. |
Similar Papers
Fair or Framed? Political Bias in News Articles Generated by LLMs (2025.emnlp-main)
Copied to clipboard
| Challenge: | Recent Large Language Models (LLMs) have garnered significant attention for applications like news generation and opinion analysis. |
| Approach: | They analyze 10,850 articles and analyze their publicViews dataset to find left-leaning bias persists in generation tasks. |
| Outcome: | The proposed model size and the PublicViews dataset show that left-leaning bias persists in generation tasks and neutral content remains rare even under balanced opinion settings. |
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts (2025.findings-acl)
Copied to clipboard
| Challenge: | Important efforts to characterize news media outlets in terms of their political bias and factuality are labor-intensive and prone to human biases. |
| Approach: | They propose a method that emulates criteria used by professional fact-checkers to assess the factuality and political bias of an entire outlet. |
| Outcome: | The proposed method improves on baselines and with multiple LLMs. |
Measuring and Mitigating Media Outlet Name Bias in Large Language Models (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing studies have explored the potential political biases of large language models, but limited attention has been devoted to the effects of media outlet names. |
| Approach: | They propose to quantify media outlet name biases in large language models and leverage this metric to develop an automated prompt optimization framework. |
| Outcome: | The proposed framework mitigates media outlet name biases, offering a scalable approach to enhancing the fairness of LLMs in news-related applications. |
Quantifying Generative Media Bias with a Corpus of Real-world and Generated News Articles (2024.findings-emnlp)
Copied to clipboard
| Challenge: | Existing studies focus on LLMs undertaking political questionnaires, which offers only limited insights into their biases and operational nuances. |
| Approach: | They propose to use a curated dataset to generate 56,700 synthetic articles using nine LLMs. |
| Outcome: | The proposed model can detect political biases using supervised models and LLMs. |
Media Source Matters More Than Content: Unveiling Political Bias in LLM-Generated Citations (2025.emnlp-main)
Copied to clipboard
| Challenge: | generative search engines rely on in-line citations as the key gateway to original webpages . a recent study shows that LLMs tend to cite left-leaning sources at higher rates compared to traditional retrieval systems . |
| Approach: | They construct a dataset of news articles labeled with left- or right-leaning stances . they find that LLMs tend to cite left-leansing sources at higher rates than traditional retrieval systems . |
| Outcome: | The proposed dataset shows that LLMs tend to cite left-leaning sources at higher rates than traditional retrieval systems. |
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said (2024.acl-long)
Copied to clipboard
| Challenge: | Existing benchmarks and measures focus on gender and racial biases, but political bias exists in LLMs and can lead to polarization and other harms in downstream applications. |
| Approach: | They propose to analyze the content and style of LLMs generated by political issues and propose a framework that can be scalable to other topics. |
| Outcome: | The proposed framework is easily scalable to other topics and is explainable. |
Assessing Reliability and Political Bias In LLMs’ Judgements of Formal and Material Inferences With Partisan Conclusions (2025.acl-long)
Copied to clipboard
| Challenge: | This paper examines the ability of LLMs to correctly label simple inferences with partisan conclusions. |
| Approach: | They develop a dataset with formal and material inferences with conclusions that favor either the political left or the political right. |
| Outcome: | The proposed models show that they are unreliable and political bias persists throughout the English and German datasets. |
Are LLMs Rational Investors? A Study on the Financial Bias in LLMs (2025.findings-acl)
Copied to clipboard
Yuhang Zhou, Yuchen Ni, Zhiheng Xi, Zhangyue Yin, Yu He, Gan Yunhui, Xiang Liu, Zhang Jian, Sen Liu, Xipeng Qiu, Yixin Cao, Guangnan Ye, Hongfeng Chai
| Challenge: | Existing studies on biases within specific domains, such as finance, remain limited. |
| Approach: | They propose a framework to detect, detect, analyze and mitigate financial biases in large language models. |
| Outcome: | The proposed framework reduces bias by 68% for the most biased model, according to key metrics. |
Investigating Bias in LLM-Based Bias Detection: Disparities between LLMs and Human Perception (2025.coling-main)
Copied to clipboard
| Challenge: | Detecting media bias is critical due to the spread of misinformation and disinformation on social media platforms. |
| Approach: | They investigate the presence and nature of bias within large language models and its consequential impact on media bias detection. |
| Outcome: | The proposed debiasing strategies include prompt engineering and model fine-tuning. |
Analyzing Political Bias in LLMs via Target-Oriented Sentiment Classification (2025.findings-acl)
Copied to clipboard
| Challenge: | Existing methods to analyze political biases rely on small-size intermediate tasks and the LLMs themselves. |
| Approach: | They propose an entropy-based inconsistency metric to encode political biases . they insert 1319 demographically and politically diverse politician names in 450 political sentences . |
| Outcome: | The proposed method combines high accuracy with a correct understanding of the candidate candidate. |