Challenge: Existing studies on automatic recognition of emotions in text have achieved promising results, but there is a shortage of resources for non-English languages, with few exceptions, like Chinese.
Approach: They propose to use a crowdsourced German emotion corpus to build a corpus similar to the English ISEAR emotion dataset.
Outcome: The proposed model performs well in German and English, but lacks the resources for non-English languages.

Similar Papers

x-enVENT: A Corpus of Event Descriptions with Experiencer-specific Emotion and Appraisal Annotations (2022.lrec-1)

Copied to clipboard

Challenge: Emotion classification is often formulated as the task to categorize texts into a predefined set of emotion classes.
Approach: They propose that a classification setup for emotion analysis should be performed in an integrated manner, including the different semantic roles that participate in an emotion episode.
Outcome: The proposed method reveals patterns in the co-occurrence of people’s emotions in interaction.
An Emotional Mess! Deciding on a Framework for Building a Dutch Emotion-Annotated Corpus (2020.lrec-1)

Copied to clipboard

Challenge: Existing frameworks for emotion recognition are limited and do not allow for categorical versus dimensional oppositions.
Approach: They propose to use the emotions joy, love, anger, sadness and fear as well as dimensional models to annotate texts from different domains and topics.
Outcome: The proposed frameworks are well-suited to annotate texts from different domains and topics, but the connotation of the labels strongly depends on the origin of the texts.
A (Psycho-)Linguistically Motivated Scheme for Annotating and Exploring Emotions in a Genre-Diverse Corpus (2022.lrec-1)

Copied to clipboard

Challenge: Using a linguistic perspective, emotion annotation is considered a difficult task because of the lack of consensus on emotional categories, the fuzziness of boundaries between them or the great variability of emotion expressions types.
Approach: They propose a scheme for emotion annotation and its manual application on a genre-diverse corpus of texts written in french.
Outcome: The proposed method clarifies the main concepts implied by the analysis of emotions as they are expressed in texts and performs a manual annotation campaign on a corpus of 1,594 texts (ca. 515K tokens) of different genres.
An Analysis of Annotated Corpora for Emotion Classification in Text (C18-1)

Copied to clipboard

Challenge: Several datasets have been annotated and published for classification of emotions.
Approach: They aggregated emotion corpora in a common file format with a shared annotation schema . they perform cross-corpus classification experiments to gain insight and a better understanding of differences .
Outcome: The proposed model can be trained on a subset of corpora, but not on all corporata.
Hard Emotion Test Evaluation Sets for Language Models (2025.findings-naacl)

Copied to clipboard

Challenge: Existing tests on emotion datasets do not show whether language models understand emotions or exploit supperficial lexical cues.
Approach: They propose to use two existing emotion datasets to evaluate whether language models make inferential decisions for emotion detection.
Outcome: The proposed test sets evaluate language models on emotion datasets.
Interpretable Relevant Emotion Ranking with Event-Driven Attention (D19-1)

Copied to clipboard

Challenge: Existing studies ignore the latent event information in documents . Existing methods for detecting emotions are limited to a few words .
Approach: They propose to integrate event information into a deep learning architecture to extract relevant emotion ranking models using corpus-level event embeddings and document-level events.
Outcome: The proposed model performs better than state-of-the-art emotion detection and multi-label approaches on three real-world corpora and interpretable results shed light on the events which trigger certain emotions.
A Method for Building a Commonsense Inference Dataset based on Basic Events (2020.emnlp-main)

Copied to clipboard

Challenge: Existing approaches to acquire commonsense are limited by the general-purpose language models.
Approach: They propose a method for building a commonsense inference dataset using crowdsourcing and automatic extraction from a corpus.
Outcome: The proposed method can solve 104k commonsense inference problems in a Japanese corpus with high accuracy, but low bias.
PO-EMO: Conceptualization, Annotation, and Modeling of Aesthetic Emotions in German and English Poetry (2020.lrec-1)

Copied to clipboard

Challenge: a new study shows that literature enables engagement in a broader range of complex and subtle emotions.
Approach: They propose to use multiple emotion labels to capture mixed emotions in poetry . they evaluate an annotation experiment with experts and crowdsourcing .
Outcome: The proposed method shows that identifying aesthetic emotions is challenging in the German subset.
Exploring Multilingual Pre-trained Language Model for Aspect-based Sentiment Analysis (2026.findings-acl)

Copied to clipboard

Challenge: Aspect-based sentiment analysis studies have focused on English datasets, but labeled data is scarce.
Approach: They propose a multilingual pre-trained language model that leverages bilingual pre-training to leverage aspects-based sentiment analysis.
Outcome: The proposed model outperforms state-of-the-art models across multiple languages.
EmoEvent: A Multilingual Emotion Corpus based on different Events (2020.lrec-1)

Copied to clipboard

Challenge: In recent years, emotion detection in text has become more popular due to its potential applications in fields such as psychology, marketing, political science, among others.
Approach: They propose to use an annotated dataset to identify emotions in tweets from different events that took place in April 2019 to validate the effectiveness of the data set.
Outcome: The proposed method is based on a multilingual emotion data set based in different events that took place in April 2019 in English and Spanish.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations