Papers by Manon Revel
Arbiters of Ambivalence: Challenges of using LLMs in No-Consensus tasks (2025.findings-acl)
Copied to clipboard
| Challenge: | LLMs are increasingly being used to replace humans in "aligning" LLM training . studies question this trend, but have found they can be more effective in ambivalent scenarios where humans disagree . |
| Approach: | They develop a “no-consensus” benchmark by curating examples that encompass a variety of a priori ambivalent scenarios. |
| Outcome: | The proposed benchmarks show that LLMs can provide nuanced assessments when generating open-ended answers, but tend to take a stance on no-consensus topics when employed as judges or debaters. |