SenSALDO: Creating a Sentiment Lexicon for Swedish (L18-1)

Copied to clipboard

Challenge: sentiment analysis has seen an explosive expansion over the last decade or so . many theoretical and methodological questions remain unanswered and resource gaps unfilled .
Approach: They develop a sentiment lexicon for written (standard) Swedish using an existing dataset . they assign a real value sentiment score in the range [-1,1] and produce a label for it .
Outcome: The proposed sentiment lexicon is an open source resource from the Swedish Language Bank . it is based on an existing gold standard dataset and is available from Sprkbanken .

Similar Papers

Generating a Gold Standard for a Swedish Sentiment Lexicon (L18-1)

Copied to clipboard

Challenge: Existing sentiment lexicons are compiled by (machine) translation from English resources, obscuring language-specific characteristics of sentiment-loaded vocabulary.
Approach: They propose a gold standard for sentiment annotation of Swedish terms using the SALDO lexicon and the Gigaword corpus.
Outcome: The proposed model is based on the free SALDO lexicon and the Gigaword corpus and is compared with existing models using human annotations.
Building Sentiment Lexicons for Mainland Scandinavian Languages Using Machine Translation and Sentence Embeddings (2022.lrec-1)

Copied to clipboard

Challenge: a simple but effective method to build sentiment lexicons for the three Mainland Scandinavian languages is proposed . a number of experiments with Scandinavian language datasets yield state-of-the-art results using a rule-based sentiment analysis algorithm.
Approach: They propose a simple but effective method to build sentiment lexicons for the three Mainland Scandinavian languages.
Outcome: The proposed method is based on the English Sentiwordnet and a thesaurus in one of the target languages.
A Thesaurus-based Sentiment Lexicon for Danish: The Danish Sentiment Lexicon (2022.lrec-1)

Copied to clipboard

Challenge: a newly published Danish sentiment lexicon with a high lexical coverage was compiled using lexicographic methods and linked data.
Approach: They propose to use lexicographic methods to compile a Danish sentiment lexicon with a high lexical coverage by linking words from a thesaurus to a comprehensive monolingual dictionary.
Outcome: The proposed lexicon contains 13,859 Danish polarity lemmas and includes morphological information.
Odi et Amo. Creating, Evaluating and Extending Sentiment Lexicons for Latin. (2020.lrec-1)

Copied to clipboard

Challenge: a new paper aims to provide sentiment analysis tools for ancient languages . the current sentiment analysis resources only cover modern languages based on textual typologies .
Approach: They propose to use manually-curated Latin lexicons to evaluate sentiment analysis tools . they propose a gold standard and a silver standard for evaluating lexical items .
Outcome: The proposed lexicons are evaluated using a gold standard and a silver standard for sentiment analysis.
Learning and Evaluating Emotion Lexicons for 91 Languages (2020.acl-main)

Copied to clipboard

Challenge: Emotion lexicons describe the affective meaning of words but are limited in coverage for most languages.
Approach: They propose a method for creating arbitrarily large emotion lexicons for any target language.
Outcome: The proposed method exceeds human reliability for some languages and variables.
Representation Mapping: A Novel Approach to Generate High-Quality Multi-Lingual Emotion Lexicons (L18-1)

Copied to clipboard

Challenge: Existing representational frameworks for emotion encoding are incompatible with semantic polarity, resulting in a large amount of incompatible emotion lexicons.
Approach: They propose to map different emotion representation formats onto each other for mutual compatibility and interoperability of language resources.
Outcome: The proposed method produces (near-)gold quality emotion lexicons even in crosslingual settings.
Encoding Sentiment Information into Word Vectors for Sentiment Analysis (C18-1)

Copied to clipboard

Challenge: Existing methods for embedding sentiment knowledge into word vectors are generally trained independently of the downstream task.
Approach: They propose to encode sentiment knowledge into pre-trained word vectors to improve sentiment analysis.
Outcome: The proposed method improves sentiment analysis on four popular sentiment datasets compared to benchmark methods.
Utilizing Large Twitter Corpora to Create Sentiment Lexica (L18-1)

Copied to clipboard

Challenge: Existing sentiment analysis systems only use word unigrams and bigrams, but lexicons using sentiment lexica are effective.
Approach: They describe an automatic Twitter sentiment lexicon creator and a lexico-based sentiment analysis system.
Outcome: The proposed system outperforms a manually annotated system in a comparison experiment.
Resource Creation Towards Automated Sentiment Analysis in Telugu (a low resource language) and Integrating Multiple Domain Sources to Enhance Sentiment Prediction (L18-1)

Copied to clipboard

Challenge: Sentiment Analysis of text is an important task in many applications . but the task becomes challenging when it comes to low resource languages .
Approach: They propose to create a corpus of polarity-based sentiment classifiers in Telugu for different domains like movie reviews, song lyrics, product reviews and book reviews.
Outcome: The proposed model performs well in multiple domains and is compared with the previous models.
Superlim: A Swedish Language Understanding Evaluation Benchmark (2023.emnlp-main)

Copied to clipboard

Challenge: In this paper, we present a multi-task benchmark for Swedish language models . we address methodological challenges, such as mitigating the Anglocentric bias when creating datasets for a less-resourced language .
Approach: They propose a multi-task NLP benchmark for Swedish language models . they propose to use superlim to evaluate Swedish language model performance .
Outcome: The proposed benchmark does not approach ceiling performance on any of the tasks, suggesting it is difficult to implement.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations