Do Emotions Influence Moral Judgment in Large Language Models? (2026.findings-acl)
Copied to clipboard
| Challenge: | Recent systems enforce explicit ethical constraints, but moral judgment rarely involves such clear-cut prohibitions. |
| Approach: | They develop an emotion-induction pipeline that infuses emotion into moral situations and evaluate shifts in moral acceptability across datasets and LLMs. |
| Outcome: | The proposed pipeline can infuses emotion into moral situations and evaluate moral acceptability shifts across datasets and LLMs. |
Similar Papers
Whose Emotions and Moral Sentiments do Language Models Reflect? (2024.findings-acl)
Copied to clipboard
| Challenge: | Existing research has focused on positional alignment, which measures how closely the models mimic the opinions and stances of different social groups. |
| Approach: | They define the problem of affective alignment, which measures how LMs’ emotional and moral tone represents those of different groups. |
| Outcome: | The results show that the models represent the perspectives of some social groups better than others, suggesting a systemic bias within LMs. |
How Inclusively do LMs Perceive Social and Moral Norms? (2025.findings-naacl)
Copied to clipboard
| Challenge: | Language models (LMs) are used in decision-making systems and as interactive assistants. |
| Approach: | They propose to prompt 11 LMs on rules-of-thumb and compare their outputs with 100 human annotators. |
| Outcome: | The proposed model is compared with 100 human annotators to find out if they are inclusive of diverse human values. |
Guilt by Association: Emotion Intensities in Lexical Representations (2021.emnlp-main)
Copied to clipboard
| Challenge: | linguistic models have a higher correlation with human ground truth ratings than labeled data . word vectors have often been evaluated on standard word relatedness benchmarks . |
| Approach: | They propose to use unsupervised, supervised, and finally supervised methods to extract emotional associations from pretrained vectors and models. |
| Outcome: | The proposed method shows higher correlation with ground truth ratings than state-of-the-art lexicons based on labeled data. |
Cross-Lingual Emotion Lexicon Induction using Representation Alignment in Low-Resource Settings (2020.coling-main)
Copied to clipboard
| Challenge: | Emotion lexicons provide information about associations between words and emotions. |
| Approach: | They use crowdsourcing to annotate words with Plutchik's 8 basic emotions, providing binary labels. |
| Outcome: | The proposed lexicons provide information about associations between words and emotions . the lexiconics are useful in emotional analyses of reviews, literary texts, and posts on social media . |
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test (2024.eacl-long)
Copied to clipboard
| Challenge: | Existing studies have shown that moral judgment depends on the language in which the dilemma is presented. |
| Approach: | They extend the work of beyond English, to 5 new languages (Chinese, Hindi, Russian, Spanish and Swahili) and probe three LLMs that show substantial multilingual text processing and generation abilities. |
| Outcome: | The models show substantial multilingual text processing and generation abilities. |
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Large language models can lead to undesired consequences when misaligned with human values . previous studies have shown misalignment of LLMs with human value using expert-designed or agent-based emulated bias scenarios . |
| Approach: | They investigate whether large language models (LLMs) are misaligned with human values . they find no significant differences in understanding of HVSB between LLMs . |
| Outcome: | The results show that large language models do not have lower misalignment rates and attack success rates . the study also shows that smaller language models have the ability to explain HVSB . |
Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation Framework (2025.emnlp-main)
Copied to clipboard
| Challenge: | Large language models (LLMs) are increasingly deployed in domains requiring moral understanding, yet their reasoning often remains shallow and misaligned with human reasoning. |
| Approach: | They propose a value-grounded framework for evaluating and distilling structured moral reasoning in large language models. |
| Outcome: | The proposed framework evaluates 12 open-source models across four moral datasets. |
Mechanistic Interpretability of Emotion Inference in Large Language Models (2025.findings-acl)
Copied to clipboard
| Challenge: | Existing studies on large language models (LLMs) show promising capabilities in predicting human emotions from text. |
| Approach: | They investigate how autoregressive LLMs infer emotions by focusing on appraisal theory . they show that emotion representations are functionally localized to specific regions in the model . |
| Outcome: | The proposed model is functionally localized to specific regions in the model, and the results align with theoretical and intuitive expectations. |
Tales of Morality: Comparing Human- and LLM-Generated Moral Stories from Visual Cues (2025.findings-emnlp)
Copied to clipboard
| Challenge: | a recent study has found that stories are central to how humans communicate moral values . |
| Approach: | They compare human- and LLM-generated moral narratives based on images annotated by humans for moral content . authors propose a framework for evaluating moral storytelling in vision-language models . |
| Outcome: | The proposed model compared human- and LLM-generated narratives on images . human stories reflect a balanced distribution of moral foundations and coherent narrative arcs, but LLMs emphasize Care foundation and lack emotional resolution. |
Feeling Rules in Language Models: Mapping Norms of Emotional Appropriateness Across Roles, Institutions, and Intensity (2026.acl-long)
Copied to clipboard
| Challenge: | Existing benchmarks measure whether Large language models recognize emotions . authors: LLMs can be used to validate, but they can still judge anger inappropriately . |
| Approach: | They propose a benchmark to measure whether Large language models validate anger . they use explicit norm judgments and implicit acceptability tests to measure norms . |
| Outcome: | The study finds that large differences in sanctioning thresholds and institutional norm signatures are not reducible to overall strictness. |