Papers by Ruben Mathew
DialDefer: A Framework for Detecting and Mitigating LLM Dialogic Deference (2026.acl-long)
Copied to clipboard
Parisa Rabbani, Priyam Sahoo, Ruben Mathew, Aishee Mondal, Harshita Ketharaman, Nimet Beyza Bozdag, Dilek Hakkani-Tür
| Challenge: | a single model can shift toward disagreement (skepticism) on graduate-level science and toward agreement (deference) on social judgment. |
| Approach: | They propose a framework to detect and mitigat framing-induced judgment shifts . they propose 'DialDefer' framework to help model disagreements and disagreements based on attribution . |
| Outcome: | The proposed framework detects and mitigates dialogic deference shifts in LLMs . human-vs-LLM attribution drives the largest shifts (17.7 pp swing) |