Papers by Fiona Liausvia
RAP: A Metric for Balancing Repetition and Performance in Open-Source Large Language Models (2025.naacl-long)
Copied to clipboard
| Challenge: | Large Language Models generate repetitive content, leading to incomplete or fragmented responses, which can negatively affect user experience. |
| Approach: | They propose a new evaluation metric that quantifies and integrates repetition penalty into the assessment of model performance, enabling tuning of RPP. |
| Outcome: | The proposed evaluation metric reduces repetition while minimizing performance loss. |