Papers by Shubhra Ghosh
Don’t Judge a Book by its Cover: Testing LLMs’ Robustness Under Logical Obfuscation (2026.eacl-long)
Copied to clipboard
| Challenge: | obfuscated questions pose significant challenges for large language models . current models parse questions without deep understanding, MIT researchers say . |
| Approach: | They propose a structure-preserving framework for logical obfuscation to test models . they use a logically equivalent framework to obliviate questions to logical equivalents . |
| Outcome: | The proposed framework is a first-of-its-kind diagnostic benchmark with 1,108 questions . obfuscation severely degrades zero-shot performance, the authors show . |