Challenge: Large language models (LLMs) are widely used in search engines to provide direct an-swers, while AI chatbots retrieve updated infor-mation from the web.
Approach: They audit nine Large Language Models from OpenAI, Google, and Meta to assess their ability to eval-uate the credibility and political bias of the top20 most popular news outlets in Bangladesh.
Outcome: The proposed models show internal consistency in credibil-ity ratings, but misalignment with human experts.

Similar Papers

Fair or Framed? Political Bias in News Articles Generated by LLMs (2025.emnlp-main)

Copied to clipboard

Challenge: Recent Large Language Models (LLMs) have garnered significant attention for applications like news generation and opinion analysis.
Approach: They analyze 10,850 articles and analyze their publicViews dataset to find left-leaning bias persists in generation tasks.
Outcome: The proposed model size and the PublicViews dataset show that left-leaning bias persists in generation tasks and neutral content remains rare even under balanced opinion settings.
Profiling News Media for Factuality and Bias Using LLMs and the Fact-Checking Methodology of Human Experts (2025.findings-acl)

Copied to clipboard

Challenge: Important efforts to characterize news media outlets in terms of their political bias and factuality are labor-intensive and prone to human biases.
Approach: They propose a method that emulates criteria used by professional fact-checkers to assess the factuality and political bias of an entire outlet.
Outcome: The proposed method improves on baselines and with multiple LLMs.
Measuring and Mitigating Media Outlet Name Bias in Large Language Models (2025.emnlp-main)

Copied to clipboard

Challenge: Existing studies have explored the potential political biases of large language models, but limited attention has been devoted to the effects of media outlet names.
Approach: They propose to quantify media outlet name biases in large language models and leverage this metric to develop an automated prompt optimization framework.
Outcome: The proposed framework mitigates media outlet name biases, offering a scalable approach to enhancing the fairness of LLMs in news-related applications.
Quantifying Generative Media Bias with a Corpus of Real-world and Generated News Articles (2024.findings-emnlp)

Copied to clipboard

Challenge: Existing studies focus on LLMs undertaking political questionnaires, which offers only limited insights into their biases and operational nuances.
Approach: They propose to use a curated dataset to generate 56,700 synthetic articles using nine LLMs.
Outcome: The proposed model can detect political biases using supervised models and LLMs.
Media Source Matters More Than Content: Unveiling Political Bias in LLM-Generated Citations (2025.emnlp-main)

Copied to clipboard

Challenge: generative search engines rely on in-line citations as the key gateway to original webpages . a recent study shows that LLMs tend to cite left-leaning sources at higher rates compared to traditional retrieval systems .
Approach: They construct a dataset of news articles labeled with left- or right-leaning stances . they find that LLMs tend to cite left-leansing sources at higher rates than traditional retrieval systems .
Outcome: The proposed dataset shows that LLMs tend to cite left-leaning sources at higher rates than traditional retrieval systems.
Measuring Political Bias in Large Language Models: What Is Said and How It Is Said (2024.acl-long)

Copied to clipboard

Challenge: Existing benchmarks and measures focus on gender and racial biases, but political bias exists in LLMs and can lead to polarization and other harms in downstream applications.
Approach: They propose to analyze the content and style of LLMs generated by political issues and propose a framework that can be scalable to other topics.
Outcome: The proposed framework is easily scalable to other topics and is explainable.
Assessing Reliability and Political Bias In LLMs’ Judgements of Formal and Material Inferences With Partisan Conclusions (2025.acl-long)

Copied to clipboard

Challenge: This paper examines the ability of LLMs to correctly label simple inferences with partisan conclusions.
Approach: They develop a dataset with formal and material inferences with conclusions that favor either the political left or the political right.
Outcome: The proposed models show that they are unreliable and political bias persists throughout the English and German datasets.
Are LLMs Rational Investors? A Study on the Financial Bias in LLMs (2025.findings-acl)

Copied to clipboard

Challenge: Existing studies on biases within specific domains, such as finance, remain limited.
Approach: They propose a framework to detect, detect, analyze and mitigate financial biases in large language models.
Outcome: The proposed framework reduces bias by 68% for the most biased model, according to key metrics.
Investigating Bias in LLM-Based Bias Detection: Disparities between LLMs and Human Perception (2025.coling-main)

Copied to clipboard

Challenge: Detecting media bias is critical due to the spread of misinformation and disinformation on social media platforms.
Approach: They investigate the presence and nature of bias within large language models and its consequential impact on media bias detection.
Outcome: The proposed debiasing strategies include prompt engineering and model fine-tuning.
Analyzing Political Bias in LLMs via Target-Oriented Sentiment Classification (2025.findings-acl)

Copied to clipboard

Challenge: Existing methods to analyze political biases rely on small-size intermediate tasks and the LLMs themselves.
Approach: They propose an entropy-based inconsistency metric to encode political biases . they insert 1319 demographically and politically diverse politician names in 450 political sentences .
Outcome: The proposed method combines high accuracy with a correct understanding of the candidate candidate.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations