MoralDial: A Framework to Train and Evaluate Moral Dialogue Systems via Moral Discussions (2023.acl-long)
Copied to clipboard
| Challenge: | A moral dialogue system aligned with users’ values could enhance conversation engagement and user connections. |
| Approach: | They propose a framework to train and evaluate moral dialogue systems based on communication mechanisms of morality and a method to construct moral discussions between simulated users and the dialogue system. |
| Outcome: | The proposed framework can train and evaluate moral dialogue systems based on simulated users and their values . |
Similar Papers
The Moral Integrity Corpus: A Benchmark for Ethical Dialogue Systems (2022.acl-long)
Copied to clipboard
| Challenge: | Moral integrity corpus captures the moral assumptions of 38k prompt-reply pairs, using 99k distinct Rules of Thumb (RoTs). |
| Approach: | They propose a resource that captures the moral assumptions of 38k prompt-reply pairs, using 99k distinct Rules of Thumb (RoTs). |
| Outcome: | The proposed resource captures the moral assumptions of 38k prompt-reply pairs, using 99k distinct Rules of Thumb (RoTs). |
A Survey on Modelling Morality for Text Analysis (2024.findings-acl)
Copied to clipboard
| Challenge: | Recent work on modelling morality in text has garnered increasing attention due to its complexity and complexity. |
| Approach: | They provide a systematic review of recent work on modelling morality in text . they discuss challenges and research gaps in the area of NLP . |
| Outcome: | The authors present their work on the modelling of morality in text, which has garnered increasing attention in recent years. |
Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation Framework (2025.emnlp-main)
Copied to clipboard
| Challenge: | Large language models (LLMs) are increasingly deployed in domains requiring moral understanding, yet their reasoning often remains shallow and misaligned with human reasoning. |
| Approach: | They propose a value-grounded framework for evaluating and distilling structured moral reasoning in large language models. |
| Outcome: | The proposed framework evaluates 12 open-source models across four moral datasets. |
The Moral Debater: A Study on the Computational Generation of Morally Framed Arguments (2022.acl-long)
Copied to clipboard
| Challenge: | Existing arguments that focus on shared values are based on prior beliefs and morals, but little research has been done on the effectiveness of these proxies. |
| Approach: | They propose a system that automatically generates arguments focusing on different morals and ask liberals and conservatives to evaluate the impact of these arguments. |
| Outcome: | The proposed system generates arguments focusing on different morals, and the results are compared with existing arguments. |
Evaluating Dialogue Generation Systems via Response Selection (2020.acl-main)
Copied to clipboard
| Challenge: | Existing automatic evaluation metrics for open-domain dialogue systems correlate poorly with human evaluation. |
| Approach: | They propose to construct response selection test sets with well-chosen false candidates to evaluate response generation systems via response selection. |
| Outcome: | The proposed method correlates with human evaluation better than widely used metrics such as BLEU. |
Rethinking Machine Ethics – Can LLMs Perform Moral Reasoning through the Lens of Moral Theories? (2024.findings-naacl)
Copied to clipboard
| Challenge: | Existing approaches to making moral judgments are mostly bottom-up and lack explainability. |
| Approach: | They propose a top-down framework to steer Large Language Models to perform moral reasoning with well-established moral theories. |
| Outcome: | The proposed framework can integrate various moral theories on moral datasets. |
Dialogue Systems for Emotional Support via Value Reinforcement (2025.acl-long)
Copied to clipboard
| Challenge: | Emotional support dialogue systems aim to reduce help-seekers’ distress and help them overcome challenges. |
| Approach: | They propose a value-driven method for training emotional support dialogue systems designed to reinforce positive values in seekers by leveraging online support conversations from Reddit. |
| Outcome: | The proposed model outperforms baseline models across support skills, seekers’ emotional intensity, and value reinforcement. |
Persuasion for Good: Towards a Personalized Persuasive Dialogue System for Social Good (P19-1)
Copied to clipboard
| Challenge: | Persuasion agents are a form of communication that can be used to change people's opinions and actions for social good. |
| Approach: | They designed an online persuasion task where one participant was asked to persult the other to donate to a specific charity. |
| Outcome: | The proposed system could change people's opinions and actions for social good. |
DialGuide: Aligning Dialogue Model Behavior with Developer Guidelines (2023.findings-emnlp)
Copied to clipboard
Prakhar Gupta, Yang Liu, Di Jin, Behnam Hedayatnia, Spandana Gella, Sijia Liu, Patrick Lange, Julia Hirschberg, Dilek Hakkani-Tur
| Challenge: | Dialogue models are able to generate fluent and interesting responses, but they can be difficult to control and may produce non-engaging, unsafe results. |
| Approach: | They propose a framework for controlling dialogue model behavior using natural language rules, or guidelines, which provide information about the context they are applicable to and what should be included in the response. |
| Outcome: | The proposed framework is effective in three open-domain dialogue response generation tasks and is consistent with the developer's expectations and intent. |
Values, Ethics, Morals? On the Use of Moral Concepts in NLP Research (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Recent studies have focused on the ethical aspects of NLP, but little to no discussion of the terminology and theories underpinning those efforts and their implications. |
| Approach: | They propose to provide an overview of some important ethical concepts stemming from philosophy and to survey the existing literature on moral NLP w.r.t. their findings show that most papers neither provide a clear definition of the terms they use nor adhere to definitions from philosophy. |
| Outcome: | The findings show that most papers neither provide a clear definition of the terms they use nor adhere to definitions from philosophy. |