Papers by Daria Kotova
What Makes Cryptic Crosswords Challenging for LLMs? (2025.coling-main)
Copied to clipboard
| Challenge: | Recent research suggests that solving cryptic crosswords is challenging even for modern NLP models, including Large Language Models (LLMs). |
| Approach: | They establish benchmark results for three popular LLMs: Gemma2, LLaMA3 and ChatGPT, and investigate why these models struggle to achieve superior performance. |
| Outcome: | The proposed models perform significantly below humans on the cryptic crossword puzzle task, while human solvers achieve 99% accuracy. |