Papers by Radu Ionescu

3 papers
RoDia: A New Dataset for Romanian Dialect Identification from Speech (2024.findings-naacl)

Copied to clipboard

Challenge: a dataset for Romanian dialect identification from speech is released . the dataset includes speech samples from five distinct regions of Romania .
Approach: They propose a dataset for Romanian dialect identification from speech . they propose competitive models to be used as baselines for future research .
Outcome: The first dataset for Romanian dialect identification from speech is released . the top scoring model achieves 59.83% and 62.08%, respectively .
A Novel Contrastive Learning Method for Clickbait Detection on RoCliCo: A Romanian Clickbait Corpus of News Articles (2023.findings-emnlp)

Copied to clipboard

Challenge: Clickbait detection is a task that aims to automatically detect misleading news titles . despite the importance of the task, there is no publicly available clickbait corpus for Romanian .
Approach: They propose a Romanian Clickbait Corpus that automatically detects misleading news titles . they propose four machine learning methods to establish competitive baselines .
Outcome: The proposed model can learn to encode news titles and contents into a deep metric space . the proposed model is available for download on github.com/dariabroscoteanu/RoCliCo.
A Novel Cartography-Based Curriculum Learning Method Applied on RoNLI: The First Romanian Natural Language Inference Corpus (2024.acl-long)

Copied to clipboard

Challenge: Natural language inference (NLI) is an actively studied topic serving as a proxy for natural language understanding.
Approach: They propose to use a Romanian NLI corpus to analyze sentence pairs . they use multiple machine learning methods to establish competitive baselines .
Outcome: The proposed model improves on the best model by employing a new curriculum learning strategy based on data cartography.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations