Crowdsourcing and Validating Event-focused Emotion Corpora for German and English (P19-1)
Copied to clipboard
| Challenge: | Existing studies on automatic recognition of emotions in text have achieved promising results, but there is a shortage of resources for non-English languages, with few exceptions, like Chinese. |
| Approach: | They propose to use a crowdsourced German emotion corpus to build a corpus similar to the English ISEAR emotion dataset. |
| Outcome: | The proposed model performs well in German and English, but lacks the resources for non-English languages. |
Similar Papers
x-enVENT: A Corpus of Event Descriptions with Experiencer-specific Emotion and Appraisal Annotations (2022.lrec-1)
Copied to clipboard
| Challenge: | Emotion classification is often formulated as the task to categorize texts into a predefined set of emotion classes. |
| Approach: | They propose that a classification setup for emotion analysis should be performed in an integrated manner, including the different semantic roles that participate in an emotion episode. |
| Outcome: | The proposed method reveals patterns in the co-occurrence of people’s emotions in interaction. |
An Emotional Mess! Deciding on a Framework for Building a Dutch Emotion-Annotated Corpus (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing frameworks for emotion recognition are limited and do not allow for categorical versus dimensional oppositions. |
| Approach: | They propose to use the emotions joy, love, anger, sadness and fear as well as dimensional models to annotate texts from different domains and topics. |
| Outcome: | The proposed frameworks are well-suited to annotate texts from different domains and topics, but the connotation of the labels strongly depends on the origin of the texts. |
A (Psycho-)Linguistically Motivated Scheme for Annotating and Exploring Emotions in a Genre-Diverse Corpus (2022.lrec-1)
Copied to clipboard
| Challenge: | Using a linguistic perspective, emotion annotation is considered a difficult task because of the lack of consensus on emotional categories, the fuzziness of boundaries between them or the great variability of emotion expressions types. |
| Approach: | They propose a scheme for emotion annotation and its manual application on a genre-diverse corpus of texts written in french. |
| Outcome: | The proposed method clarifies the main concepts implied by the analysis of emotions as they are expressed in texts and performs a manual annotation campaign on a corpus of 1,594 texts (ca. 515K tokens) of different genres. |
An Analysis of Annotated Corpora for Emotion Classification in Text (C18-1)
Copied to clipboard
| Challenge: | Several datasets have been annotated and published for classification of emotions. |
| Approach: | They aggregated emotion corpora in a common file format with a shared annotation schema . they perform cross-corpus classification experiments to gain insight and a better understanding of differences . |
| Outcome: | The proposed model can be trained on a subset of corpora, but not on all corporata. |
Hard Emotion Test Evaluation Sets for Language Models (2025.findings-naacl)
Copied to clipboard
| Challenge: | Existing tests on emotion datasets do not show whether language models understand emotions or exploit supperficial lexical cues. |
| Approach: | They propose to use two existing emotion datasets to evaluate whether language models make inferential decisions for emotion detection. |
| Outcome: | The proposed test sets evaluate language models on emotion datasets. |
Interpretable Relevant Emotion Ranking with Event-Driven Attention (D19-1)
Copied to clipboard
| Challenge: | Existing studies ignore the latent event information in documents . Existing methods for detecting emotions are limited to a few words . |
| Approach: | They propose to integrate event information into a deep learning architecture to extract relevant emotion ranking models using corpus-level event embeddings and document-level events. |
| Outcome: | The proposed model performs better than state-of-the-art emotion detection and multi-label approaches on three real-world corpora and interpretable results shed light on the events which trigger certain emotions. |
A Method for Building a Commonsense Inference Dataset based on Basic Events (2020.emnlp-main)
Copied to clipboard
| Challenge: | Existing approaches to acquire commonsense are limited by the general-purpose language models. |
| Approach: | They propose a method for building a commonsense inference dataset using crowdsourcing and automatic extraction from a corpus. |
| Outcome: | The proposed method can solve 104k commonsense inference problems in a Japanese corpus with high accuracy, but low bias. |
PO-EMO: Conceptualization, Annotation, and Modeling of Aesthetic Emotions in German and English Poetry (2020.lrec-1)
Copied to clipboard
| Challenge: | a new study shows that literature enables engagement in a broader range of complex and subtle emotions. |
| Approach: | They propose to use multiple emotion labels to capture mixed emotions in poetry . they evaluate an annotation experiment with experts and crowdsourcing . |
| Outcome: | The proposed method shows that identifying aesthetic emotions is challenging in the German subset. |
Exploring Multilingual Pre-trained Language Model for Aspect-based Sentiment Analysis (2026.findings-acl)
Copied to clipboard
| Challenge: | Aspect-based sentiment analysis studies have focused on English datasets, but labeled data is scarce. |
| Approach: | They propose a multilingual pre-trained language model that leverages bilingual pre-training to leverage aspects-based sentiment analysis. |
| Outcome: | The proposed model outperforms state-of-the-art models across multiple languages. |
EmoEvent: A Multilingual Emotion Corpus based on different Events (2020.lrec-1)
Copied to clipboard
| Challenge: | In recent years, emotion detection in text has become more popular due to its potential applications in fields such as psychology, marketing, political science, among others. |
| Approach: | They propose to use an annotated dataset to identify emotions in tweets from different events that took place in April 2019 to validate the effectiveness of the data set. |
| Outcome: | The proposed method is based on a multilingual emotion data set based in different events that took place in April 2019 in English and Spanish. |