Papers by Ningxin Wu
Beyond Detection: Evaluating Fallacy Awareness of LLMs in Interactive Scenarios (2026.acl-long)
Copied to clipboard
| Challenge: | Large Language Models fail to recognize fallacious reasoning in real-world interactions despite strong performance on static fallacy detection tasks. |
| Approach: | They propose a Chinese benchmark to assess fallacy awareness without explicit cues . they propose 'fate' evaluation framework that assesses fallacy without explicit . |
| Outcome: | The proposed framework assesses fallacy awareness without explicit cues, combining natural dialogue responses and reasoning-based decisions. |