Papers by María-Teresa Martín-Valdivia
SHARE: A Lexicon of Harmful Expressions by Spanish Speakers (2022.lrec-1)
Copied to clipboard
Flor Miriam Plaza-del-Arco, Ana Belén Parras Portillo, Pilar López Úbeda, Beatriz Gil, María-Teresa Martín-Valdivia
| Challenge: | Using natural language processing, offensive comments can be created by composition of words. |
| Approach: | They propose to use a lexical resource with 10,125 offensive terms and expressions collected from Spanish speakers to retrieve the vocabulary. |
| Outcome: | The proposed resource has 10,125 offensive terms and expressions and is used to identify spans in Spanish. |
La Leaderboard: A Large Language Model Leaderboard for Spanish Varieties and Languages of Spain and Latin America (2025.acl-long)
Copied to clipboard
María Grandury, Javier Aula-Blasco, Júlia Falcão, Clémentine Fourrier, Miguel González Saiz, Gonzalo Martínez, Gonzalo Santamaria Gomez, Rodrigo Agerri, Nuria Aldama García, Luis Chiruzzo, Javier Conde, Helena Gomez Adorno, Marta Guerrero Nieto, Guido Ivetta, Natàlia López Fuertes, Flor Miriam Plaza-del-Arco, María-Teresa Martín-Valdivia, Helena Montoro Zamorano, Carmen Muñoz Sanz, Pedro Reviriego, Leire Rosado Plaza, Alejandro Vaca Serrano, Estrella Vallecillo-Rodríguez, Jorge Vallego, Irune Zubiaga
| Challenge: | La Leaderboard is the first open-source leaderboard to evaluate generative Large Language Models (LLMs) in languages and language varieties of Spain and Latin America. |
| Approach: | They propose to use La Leaderboard to evaluate generative Large Language Models in Spanish and Latin America. |
| Outcome: | La Leaderboard is the first open-source leaderboard to evaluate generative LLMs in languages and language varieties of Spain and Latin America. |
Natural Language Inference Prompts for Zero-shot Emotion Classification in Text across Corpora (2022.coling-1)
Copied to clipboard
| Challenge: | Existing models for textual emotion classification depend on domain and application scenario and need to be predefined . a natural language inference model with a flexible set of labels is difficult to develop . |
| Approach: | They propose to use the paradigm of zero-shot learning as a natural language inference task to generate a model with a flexible set of labels. |
| Outcome: | The proposed model is more robust across corpora than individual prompts and shows similar performance to the best prompt for a particular corpus. |
Nuanced Toxicity Detection in Spanish: A New Corpus and Benchmark Study (2026.findings-eacl)
Copied to clipboard
Alba María Mármol-Romero, Robiert Sepúlveda-Torres, Estela Saquete, María-Teresa Martín-Valdivia, L. Alfonso Ureña
| Challenge: | Existing corpora for Spanish are under-resourced for toxic content detection . sarcasm, indirect aggression, irony, and other toxicity are not detected in English . |
| Approach: | They propose to extend the NECOS-TOX corpus to include 4,011 Spanish comments . each comment is annotated across three levels of toxicity, with substantial inter-annotator agreement . |
| Outcome: | The proposed model performs on par with larger models and is released publicly . the proposed model is based on a human-in-the-loop active learning strategy . |