Papers by Johanna Ramirez-Romero
Ranking Over Scoring: Towards Reliable and Robust Automated Evaluation of LLM-Generated Medical Explanatory Arguments (2025.coling-main)
Copied to clipboard
Iker De la Iglesia, Iakes Goenaga, Johanna Ramirez-Romero, Jose Maria Villa-Gonzalez, Josu Goikoetxea, Ander Barrena
| Challenge: | Evaluating LLM-generated text has become a key challenge in domain-specific contexts like the medical field. |
| Approach: | They propose a method to evaluate LLM-generated medical explanatory arguments using Proxy Tasks and rankings to align results with human evaluation criteria. |
| Outcome: | The proposed evaluation method is robust against adversarial attacks, including the assessment of non-argumentative text. |