Challenge: Large language models exhibit cultural bias from over-represented viewpoints in training data, yet cultural alignment remains a challenge due to limited cultural knowledge and a lack of exploration into effective learning approaches.
Approach: They propose a cost-efficient method for fine-tuning large language models on native speakers’ word-association norms and a preference optimization method to improve cultural alignment.
Outcome: The proposed model trains Llama-3.1-8B and Qwen-2.5-7B on native speakers’ word-association norms and shows that such associations capture cultural knowledge.

Similar Papers

From Word to World: Evaluate and Mitigate Culture Bias in LLMs via Word Association Test (2025.emnlp-main)

Copied to clipboard

Challenge: Multilingual and cross-cultural WAT reveal how culture modulates perceptual and interactive patterns.
Approach: They propose to embed cultural-specific semantic associations directly within large language models (LLMs) to address cultural preference.
Outcome: The proposed model significantly improves cross-cultural alignment, capturing diverse semantic associations.
Pedagogical Alignment of Large Language Models (2024.findings-emnlp)

Copied to clipboard

Challenge: Large Language Models (LLMs) are often used without pedagogical fine-tuning and provide immediate answers rather than guiding students through the problem-solving process.
Approach: They propose a method for constructing large-scale preference datasets using synthetic data generation techniques that eliminates the need for manual annotation.
Outcome: The proposed methods outperform standard supervised fine-tuning (SFT) and improve alignment accuracy by 13.1% and 8.7% respectively.
Investigating Cultural Alignment of Large Language Models (2024.acl-long)

Copied to clipboard

Challenge: Large Language Models (LLMs) are used to represent the diversity of human experience and culturally sensitive topics.
Approach: They propose a method leveraging anthropological reasoning to enhance cultural alignment by prompting LLMs with different pretraining data mixtures in Arabic and English.
Outcome: The proposed method enables users to better represent the diversity of human experience and the plurality of different cultures.
AlignCultura: Towards Culturally Aligned Large Language Models? (2026.acl-long)

Copied to clipboard

Challenge: Existing benchmarks represent early steps toward cultural alignment, yet no benchmarks currently enables systematic evaluation of cultural alignment in line with UNESCO’s principles of cultural diversity w.r.t HHH paradigm.
Approach: Align-Cultura aims to evaluate cultural alignment in large language models . it uses a Query Construction pipeline to reclassify prompts and expand underrepresented domains . response generation pairs prompts with culturally grounded responses .
Outcome: Empirically, culturally fine-tuned models improve joint HHH by 4%–6%, reduce cultural failures by 18%, achieve 10%–12% efficiency gains, and limit leakage to 0.3%.
Can Persona-Prompted LLMs Emulate Subgroup Values? An Empirical Analysis of Generalisability and Fairness in Cultural Alignment (2026.acl-long)

Copied to clipboard

Challenge: Current alignment paradigms treat "human values" as a monolithic entity, ignoring the fact that many societies are a mosaic of diverse subgroups with distinct and sometimes conflicting values, preferences, and norms.
Approach: They examine whether Large Language Models can emulate distinct cultural values of subgroups . they use a global value survey to examine the value landscape of a multicultural society .
Outcome: The proposed model improves on unseen, out-of-distribution subgroups by 17.4% . the model widens the disparity between subgroup groups when measured by distance-aware metrics.
Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede’s Cultural Dimensions (2025.coling-main)

Copied to clipboard

Challenge: Large language models (LLMs) are deployed in many countries, but they fail to account for cultural variances among their potential users.
Approach: They propose to use Hofstede’s cultural dimension framework to quantify cultural alignment using latent variable analysis to evaluate large language models against cultural dimensions of regions like the United States, China, and Arab countries.
Outcome: The proposed model is compared against LLMs in the United States, China, and Arab countries and demonstrates that all models struggle to grasp cultural values, while GPT-4 shows a unique capability to adapt to cultural nuances, particularly in Chinese settings.
Aligning LLMs for Multilingual Consistency in Enterprise Applications (2025.emnlp-industry)

Copied to clipboard

Challenge: Large language models (LLMs) remain unreliable for global enterprise applications due to performance gaps between high-resource and mid/low-resourced languages .
Approach: They propose a batch-wise alignment strategy that aligns model outputs across languages . this method improves non-English accuracy by up to 23.9% without compromising English performance .
Outcome: The proposed approach improves non-English accuracy by up to 23.9% without compromising English performance, model reasoning, or retrieval quality.
A Survey on Training-free Alignment of Large Language Models (2025.findings-emnlp)

Copied to clipboard

Challenge: a survey of large language models (LLMs) aims to ensure outputs adhere to human values, ethical standards, and legal norms.
Approach: They present the first systematic review of TF alignment methods . they categorize them by stages of pre-decoding, in-decoder and post-decoration .
Outcome: The proposed methods are based on training-free (TF) alignment techniques . they are able to be used in open-source and closed-source environments without retraining .
LIONs: An Empirically Optimized Approach to Align Language Models (2024.emnlp-main)

Copied to clipboard

Challenge: Recent studies have focused on aligning large language models with pre-trained datasets.
Approach: They conduct a rigorous analysis of a three-stage training pipeline using sequence packing, loss masking and increasing the preference dataset size in DPO to improve the performance of language models.
Outcome: The proposed models outperform the official instruct models tuned with closed-source data and algorithms.
Incorporating Diverse Perspectives in Cultural Alignment: Survey of Evaluation Benchmarks Through A Three-Dimensional Framework (2025.emnlp-main)

Copied to clipboard

Challenge: Large Language Models (LLMs) serve diverse global audiences, making it critical for responsible AI deployment across cultures.
Approach: They propose a framework that conceptualizes alignment along three dimensions: Cultural Group, Cultural Elements and Awareness Scope.
Outcome: The proposed framework reveals critical gaps between benchmarks and real-world cultural biases . region dominates cultural group representation, social and political relations dominates coverage . majority of datasets adopt majority-focused Awareness Scope approaches .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations