Papers by Mario Mina
Cognitive Biases, Task Complexity, and Result Intepretability in Large Language Models (2025.coling-main)
Copied to clipboard
| Challenge: | Recent work shows that cognitive biases occur frequently in language models . a cognitive bias is a systematic deviation in judgment that simplifies complex decisions . |
| Approach: | They evaluate the performance of different groups of models for each type of cognitive bias . they find that task complexity plays a part in eliciting stronger effects for some biases . |
| Outcome: | The proposed models perform better for each type of bias in different settings . the results show that task complexity plays a part in eliciting stronger effects . |
A CURATEd CATalog: Rethinking the Extraction of Pretraining Corpora for Mid-Resourced Languages (2024.lrec-main)
Copied to clipboard
Jorge Palomar-Giner, Jose Javier Saiz, Ferran Espuña, Mario Mina, Severino Da Dalt, Joan Llop, Malte Ostendorff, Pedro Ortiz Suarez, Georg Rehm, Aitor Gonzalez-Agirre, Marta Villegas
| Challenge: | CATalog 1.0 is the largest text corpus in Catalan to date . CURATE is a pipeline that can be parallelizable to run in high performance clusters . |
| Approach: | They propose a data pipeline that uses binary filters to filter documents based on text quality . they optimised the pipeline to run in high performance clusters . |
| Outcome: | The proposed pipeline is optimized for high performance cluster environments and runs in high performance. |