Papers by Iulia Comșa
MiQA: A Benchmark for Inference on Metaphorical Questions (2022.aacl-short)
Copied to clipboard
| Challenge: | a benchmark is proposed to assess the capability of large language models to reason with conventional metaphors. |
| Approach: | They propose to assess the capability of large language models to reason with conventional metaphors. |
| Outcome: | The proposed benchmark compares pre-trained models on binary-choice tasks with human models . the results show that human models perform better on the largest model, compared to small models based on the same task . |