Toward Qualitative Evaluation of Embeddings for Arabic Sentiment Analysis (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing studies on Arabic sentiment analysis (SA) tasks focus on word embeddings to capture semantic and syntactic similarities, but Arabic language is characterized by its agglutination and morphological richness contributing to great sparsity. |
| Approach: | They propose several protocols to evaluate specific embeddings for Arabic sentiment analysis task. |
| Outcome: | The proposed embeddings are based on words and lemmas in Arabic sentiment analysis (SA) task. |
Similar Papers
A Comprehensive Survey of Contemporary Arabic Sentiment Analysis: Methods, Challenges, and Future Directions (2025.findings-naacl)
Copied to clipboard
| Challenge: | Existing literature on Arabic sentiment analysis is limited, compared to high-resourced languages such as English and French. |
| Approach: | They present a systematic review of existing literature on Arabic sentiment analysis focusing on research utilizing deep learning. |
| Outcome: | The proposed methods highlight gaps in the literature on Arabic sentiment analysis and outline promising directions for future research. |
Swan and ArabicMTEB: Dialect-Aware, Arabic-Centric, Cross-Lingual, and Cross-Cultural Embedding Models and Benchmarks (2025.findings-naacl)
Copied to clipboard
Gagan Bhatia, El Moatez Billah Nagoudi, Abdellah El Mekki, Fakhraddin Alwajih, Muhammad Abdul-Mageed
| Challenge: | In this paper, we introduce a family of embedding models addressing both small-scale and large-scale use cases. |
| Approach: | They propose to use ArabicMTEB to evaluate Arabic text embedding models . they propose to build a benchmark suite that assesses cross-lingual, multi-dialectal, multidomain, and multi-cultural Arabic text embedded models. |
| Outcome: | The proposed models outperform Multilingual-E5-large and Swan-Large in most Arabic tasks while remaining dialectally and culturally aware. |
SentiArabic: A Sentiment Analyzer for Standard Arabic (L18-1)
Copied to clipboard
| Challenge: | Sentiment analysis is a process of applying computational approaches to identify attitudes, emotions and opinions in text, speech and visual data. |
| Approach: | They propose a sentiment analyzer that identifies the overall contextual polarity for Arabic text. |
| Outcome: | The proposed system achieves an F-score of 76.5% when evaluated on a blind test set. |
Towards Qualitative Word Embeddings Evaluation: Measuring Neighbors Variation (N18-4)
Copied to clipboard
| Challenge: | Using extrinsic evaluation methods, embeddings are evaluated on a specific task such as part-of-speech tagging or named-entity recognition. |
| Approach: | They propose a method to study the variation between word embeddings models trained with only one parameter by observing the distributional neighbors variation. |
| Outcome: | The proposed method shows that changing only one parameter can have a massive impact on a given semantic space. |
Domain Adaptation for Arabic Cross-Domain and Cross-Dialect Sentiment Analysis from Contextualized Word Embedding (2021.naacl-main)
Copied to clipboard
| Challenge: | Recent studies have classified dialectal Arabic into more fine-grained levels, including countries and cities. |
| Approach: | They propose to use Arabic domains to transfer knowledge from labeled source domains into unlabeled target domains by transferring the learned knowledge from a labele . |
| Outcome: | The proposed method outperforms other domain adaptation methods and improves performance by 20.8% over the zero-shot transfer learning from BERT. |
Weakly-Supervised Aspect-Based Sentiment Analysis via Joint Aspect-Sentiment Topic Embedding (2020.emnlp-main)
Copied to clipboard
| Challenge: | Existing methods for aspect-based sentiment analysis of review text use only a few keywords describing each aspect/sentiment without using any labeled examples. |
| Approach: | They propose a weakly-supervised approach for aspect-based sentiment analysis which uses only a few keywords describing each aspect/sentiment without using any labeled examples. |
| Outcome: | The proposed method generates quality joint topics and outperforms baselines significantly on benchmark datasets. |
Revisiting Pre-trained Language Models and their Evaluation for Arabic Natural Language Processing (2022.emnlp-main)
Copied to clipboard
Abbas Ghaddar, Yimeng Wu, Sunyam Bagga, Ahmad Rashid, Khalil Bibi, Mehdi Rezagholizadeh, Chao Xing, Yasheng Wang, Xinyu Duan, Zhefeng Wang, Baoxing Huai, Xin Jiang, Qun Liu, Phillippe Langlais
| Challenge: | Existing pre-trained language models are not well-explored and are not reproducible in the literature. |
| Approach: | They propose to improve existing Arabic language pre-trained language models using a more methodical approach. |
| Outcome: | The proposed models outperform existing models on ALUE, a leaderboard-powered benchmark for Arabic NLU and NLG tasks. |
Metaphorical Expressions in Automatic Arabic Sentiment Analysis (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing algorithms and tools for sentiment analysis are lacking in dealing with Arabic metaphorical expressions. |
| Approach: | They propose to use Arabic metaphors in automatic Arabic sentiment analysis to examine the performance of a state-of-art Arabic sentiment tool on metaphors. |
| Outcome: | The proposed model outperforms the state-of-the-art sentiment analysis tool on metaphors and gain a deeper insight into the issue. |
Searching for the X-Factor: Exploring Corpus Subjectivity for Word Embeddings (P18-1)
Copied to clipboard
| Challenge: | Existing word embedding methods for natural language processing are limited in their ability to produce dense word embeds. |
| Approach: | They propose a word embedding SentiVec which is infused with sentiment information from a lexical resource and outperforms baselines on subjectivity-sensitive tasks. |
| Outcome: | The proposed word embedding SentiVec outperforms baselines on subjectivity-sensitive tasks. |
DiaSet: An Annotated Dataset of Arabic Conversations (2024.lrec-main)
Copied to clipboard
Abraham Israeli, Aviv Naaman, Guy Maduel, Rawaa Makhoul, Dana Qaraeen, Amir Ejmail, Dina Lisnanskey, Julian Jubran, Shai Fine, Kfir Bar
| Challenge: | DiaSet is a dataset of dialectical Arabic speech manually transcribed and annotated for two downstream tasks. |
| Approach: | They propose to manually transcribe and annotate Arabic speech for sentiment analysis and named entity recognition. |
| Outcome: | The proposed dataset encapsulates the Palestine dialect, predominantly spoken in Palestine, Israel, and Jordan. |