Papers by Johanna Ramirez-Romero

1 papers
Ranking Over Scoring: Towards Reliable and Robust Automated Evaluation of LLM-Generated Medical Explanatory Arguments (2025.coling-main)

Copied to clipboard

Challenge: Evaluating LLM-generated text has become a key challenge in domain-specific contexts like the medical field.
Approach: They propose a method to evaluate LLM-generated medical explanatory arguments using Proxy Tasks and rankings to align results with human evaluation criteria.
Outcome: The proposed evaluation method is robust against adversarial attacks, including the assessment of non-argumentative text.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations