Papers by Gaman Mihaela
SaRoCo: Detecting Satire in a Novel Romanian Corpus of News Articles (2021.acl-short)
Copied to clipboard
| Challenge: | a corpus for satire detection in Romanian news is based on satirical reporting . the goal is to ridicule public figures, politics or contemporary events . |
| Approach: | They propose a corpus for satire detection in Romanian news . they gather 55,608 public news articles from multiple real and satirical sources . |
| Outcome: | The proposed corpus is one of the largest corpora for satire detection regardless of language . it is the only one for the Romanian language, and the results show that it is low on the machine level compared to human level . |
Automatically Identifying Complaints in Social Media (P19-1)
Copied to clipboard
| Challenge: | Complaining is a basic speech act used to express a negative mismatch between reality and expectations in a particular situation. |
| Approach: | They present a systematic analysis of complaints in computational linguistics . they collect annotated data set of written complaints expressed on Twitter . |
| Outcome: | The proposed model achieves predictive performance of up to 79 F1 using distant supervision. |
Clustering Word Embeddings with Self-Organizing Maps. Application on LaRoSeDa - A Large Romanian Sentiment Data Set (2021.eacl-main)
Copied to clipboard
| Challenge: | Romanian is one of the understudied languages in computational linguistics, with few resources available for the development of natural language processing tools. |
| Approach: | They introduce a Large Romanian Sentiment Data Set which is composed of 15,000 positive and negative reviews collected from the largest Romanian e-commerce platform. |
| Outcome: | The proposed data set is composed of 15,000 positive and negative reviews from the largest Romanian e-commerce platform. |