Challenge: Recent systems enforce explicit ethical constraints, but moral judgment rarely involves such clear-cut prohibitions.
Approach: They develop an emotion-induction pipeline that infuses emotion into moral situations and evaluate shifts in moral acceptability across datasets and LLMs.
Outcome: The proposed pipeline can infuses emotion into moral situations and evaluate moral acceptability shifts across datasets and LLMs.

Similar Papers

Whose Emotions and Moral Sentiments do Language Models Reflect? (2024.findings-acl)

Copied to clipboard

Challenge: Existing research has focused on positional alignment, which measures how closely the models mimic the opinions and stances of different social groups.
Approach: They define the problem of affective alignment, which measures how LMs’ emotional and moral tone represents those of different groups.
Outcome: The results show that the models represent the perspectives of some social groups better than others, suggesting a systemic bias within LMs.
How Inclusively do LMs Perceive Social and Moral Norms? (2025.findings-naacl)

Copied to clipboard

Challenge: Language models (LMs) are used in decision-making systems and as interactive assistants.
Approach: They propose to prompt 11 LMs on rules-of-thumb and compare their outputs with 100 human annotators.
Outcome: The proposed model is compared with 100 human annotators to find out if they are inclusive of diverse human values.
Guilt by Association: Emotion Intensities in Lexical Representations (2021.emnlp-main)

Copied to clipboard

Challenge: linguistic models have a higher correlation with human ground truth ratings than labeled data . word vectors have often been evaluated on standard word relatedness benchmarks .
Approach: They propose to use unsupervised, supervised, and finally supervised methods to extract emotional associations from pretrained vectors and models.
Outcome: The proposed method shows higher correlation with ground truth ratings than state-of-the-art lexicons based on labeled data.
Cross-Lingual Emotion Lexicon Induction using Representation Alignment in Low-Resource Settings (2020.coling-main)

Copied to clipboard

Challenge: Emotion lexicons provide information about associations between words and emotions.
Approach: They use crowdsourcing to annotate words with Plutchik's 8 basic emotions, providing binary labels.
Outcome: The proposed lexicons provide information about associations between words and emotions . the lexiconics are useful in emotional analyses of reviews, literary texts, and posts on social media .
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test (2024.eacl-long)

Copied to clipboard

Challenge: Existing studies have shown that moral judgment depends on the language in which the dilemma is presented.
Approach: They extend the work of beyond English, to 5 new languages (Chinese, Hindi, Russian, Spanish and Swahili) and probe three LLMs that show substantial multilingual text processing and generation abilities.
Outcome: The models show substantial multilingual text processing and generation abilities.
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs (2025.findings-emnlp)

Copied to clipboard

Challenge: Large language models can lead to undesired consequences when misaligned with human values . previous studies have shown misalignment of LLMs with human value using expert-designed or agent-based emulated bias scenarios .
Approach: They investigate whether large language models (LLMs) are misaligned with human values . they find no significant differences in understanding of HVSB between LLMs .
Outcome: The results show that large language models do not have lower misalignment rates and attack success rates . the study also shows that smaller language models have the ability to explain HVSB .
Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation Framework (2025.emnlp-main)

Copied to clipboard

Challenge: Large language models (LLMs) are increasingly deployed in domains requiring moral understanding, yet their reasoning often remains shallow and misaligned with human reasoning.
Approach: They propose a value-grounded framework for evaluating and distilling structured moral reasoning in large language models.
Outcome: The proposed framework evaluates 12 open-source models across four moral datasets.
Mechanistic Interpretability of Emotion Inference in Large Language Models (2025.findings-acl)

Copied to clipboard

Challenge: Existing studies on large language models (LLMs) show promising capabilities in predicting human emotions from text.
Approach: They investigate how autoregressive LLMs infer emotions by focusing on appraisal theory . they show that emotion representations are functionally localized to specific regions in the model .
Outcome: The proposed model is functionally localized to specific regions in the model, and the results align with theoretical and intuitive expectations.
Tales of Morality: Comparing Human- and LLM-Generated Moral Stories from Visual Cues (2025.findings-emnlp)

Copied to clipboard

Challenge: a recent study has found that stories are central to how humans communicate moral values .
Approach: They compare human- and LLM-generated moral narratives based on images annotated by humans for moral content . authors propose a framework for evaluating moral storytelling in vision-language models .
Outcome: The proposed model compared human- and LLM-generated narratives on images . human stories reflect a balanced distribution of moral foundations and coherent narrative arcs, but LLMs emphasize Care foundation and lack emotional resolution.
Feeling Rules in Language Models: Mapping Norms of Emotional Appropriateness Across Roles, Institutions, and Intensity (2026.acl-long)

Copied to clipboard

Challenge: Existing benchmarks measure whether Large language models recognize emotions . authors: LLMs can be used to validate, but they can still judge anger inappropriately .
Approach: They propose a benchmark to measure whether Large language models validate anger . they use explicit norm judgments and implicit acceptability tests to measure norms .
Outcome: The study finds that large differences in sanctioning thresholds and institutional norm signatures are not reducible to overall strictness.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations