Papers by Maia Aguirre
BaSCo: An Annotated Basque-Spanish Code-Switching Corpus for Natural Language Understanding (2022.lrec-1)
Copied to clipboard
| Challenge: | Basque-Spanish code-switching is a widespread phenomenon among bilingual speakers in the Basque Country. |
| Approach: | They propose to use annotated utterances to train bilingual chatbots in Basque and Spanish to cover the phenomenon of code-switching. |
| Outcome: | The proposed corpus is the first with annotated linguistic resources encompassing Basque-Spanish code-switching. |
Fine-Tuning Medium-Scale LLMs for Joint Intent Classification and Slot Filling: A Data-Efficient and Cost-Effective Solution for SMEs (2025.coling-industry)
Copied to clipboard
| Challenge: | Current techniques for user comprehension in DS depend heavily on labeled data and the data annotation process for NLU is labor-intensive and requires expert annotators. |
| Approach: | They propose to fine-tune a model for joint Intent Classification and Slot Filling with only 10% of the data. |
| Outcome: | The proposed model outperforms existing models in monolingual and cross-lingual scenarios with only 10% of the data. |
Exploiting In-Domain Bilingual Corpora for Zero-Shot Transfer Learning in NLU of Intra-Sentential Code-Switching Chatbot Interactions (2022.emnlp-industry)
Copied to clipboard
Maia Aguirre, Manex Serras, Laura García-sardiña, Jacobo López-fernández, Ariane Méndez, Arantza Del Pozo
| Challenge: | Multilingual speakers outnumber monolingual speakers in the world . CS is a frequent habit in both spoken and written informal communications . |
| Approach: | They evaluate the efficacy of cross-lingual transfer learning with mBERT for NLU on a Basque-Spanish CS chatbot corpus. |
| Outcome: | The proposed model outperforms models trained on Basque and Spanish without CS on a basque-Spanish chatbot corpus. |