Papers by Radu Ionescu
RoDia: A New Dataset for Romanian Dialect Identification from Speech (2024.findings-naacl)
Copied to clipboard
| Challenge: | a dataset for Romanian dialect identification from speech is released . the dataset includes speech samples from five distinct regions of Romania . |
| Approach: | They propose a dataset for Romanian dialect identification from speech . they propose competitive models to be used as baselines for future research . |
| Outcome: | The first dataset for Romanian dialect identification from speech is released . the top scoring model achieves 59.83% and 62.08%, respectively . |
A Novel Contrastive Learning Method for Clickbait Detection on RoCliCo: A Romanian Clickbait Corpus of News Articles (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Clickbait detection is a task that aims to automatically detect misleading news titles . despite the importance of the task, there is no publicly available clickbait corpus for Romanian . |
| Approach: | They propose a Romanian Clickbait Corpus that automatically detects misleading news titles . they propose four machine learning methods to establish competitive baselines . |
| Outcome: | The proposed model can learn to encode news titles and contents into a deep metric space . the proposed model is available for download on github.com/dariabroscoteanu/RoCliCo. |
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus (2024.acl-long)
Copied to clipboard
| Challenge: | Natural language inference (NLI) is an actively studied topic serving as a proxy for natural language understanding. |
| Approach: | They propose to use a Romanian NLI corpus to analyze sentence pairs . they use multiple machine learning methods to establish competitive baselines . |
| Outcome: | The proposed model improves on the best model by employing a new curriculum learning strategy based on data cartography. |