Papers by Muhammad Ali

10 papers
BRIGHTER: BRIdging the Gap in Human-Annotated Textual Emotion Recognition Datasets for 28 Languages (2025.acl-long)

Copied to clipboard

Challenge: Emotion recognition is an umbrella term for several NLP tasks, but most work on high-resource languages has focused on low-resourced languages.
Approach: They propose to use emotion recognition to describe perceived emotions in 28 different languages and across several domains to identify and annotate the datasets.
Outcome: The proposed datasets cover low-resource languages from Africa, Asia, Eastern Europe, and Latin America, with instances labeled by fluent speakers.
POLAR: A Benchmark for Multilingual, Multicultural, and Multi-Event Online Polarization (2026.findings-acl)

Copied to clipboard

Challenge: polarization is a pervasive threat to democratic institutions, civil discourse, and social cohesion worldwide . most existing datasets focus on English or high-resource languages, reflecting a widespread trend across NLP tasks .
Approach: They propose a multilingual, multicultural, and multi-event dataset with over 110K instances in 22 languages drawn from diverse online platforms and real-world events.
Outcome: The proposed dataset analyzes polarization detection, type, and manifestation using a variety of annotation platforms adapted to each cultural context.
NusaCrowd: Open Source Initiative for Indonesian NLP Resources (2023.findings-acl)

Copied to clipboard

Challenge: Existing NLP research in Indonesian languages has been held back by factors such as language diversity, orthographic variation, resource limitation and other societal challenges.
Approach: They present a collaborative initiative to collect and unify existing resources for Indonesian languages and open access to previously non-public resources.
Outcome: The results show that the datasets are highly reliable and can be used to generate the first zero-shot benchmarks for natural language understanding and generation in Indonesian and the local languages of Indonesia.
Beyond Content: How Grammatical Gender Shapes Visual Representation in Text-to-Image Models (2025.findings-emnlp)

Copied to clipboard

Challenge: grammatical gender significantly influences image generation in text-to-image models . masculine grammatikal markers increase male representation to 73% on average . feminine grammatological markers increase female representation to 38% .
Approach: They propose a cross-linguistic benchmark examining words where grammatical gender contradicts stereotypical gender associations.
Outcome: The proposed benchmark examines words where grammatical gender contradicts stereotypical gender associations.
AfriSenti: A Twitter Sentiment Analysis Benchmark for African Languages (2023.emnlp-main)

Copied to clipboard

Challenge: Africa has the highest linguistic diversity among all continents.
Approach: They introduce a sentiment analysis benchmark that contains >110,000 tweets in 14 African languages . they describe the data collection methodology, annotation process, and challenges .
Outcome: The proposed dataset contains >110,000 tweets in 14 African languages . the tweets were annotated by native speakers and used in the shared task .
GRI: Graph-based Relative Isomorphism of Word Embedding Spaces (2023.findings-emnlp)

Copied to clipboard

Challenge: Existing attempts to control relative isomorphism of different spaces fail to consider lexical variations of semantically similar words . Existing methods for building bilingual dictionaries rely on geometric similarity of individual spaces .
Approach: They propose a method that incorporates the impact of lexical variations of semantically similar words into the training objective.
Outcome: The proposed method outperforms existing research by improving the average P@1 by 63.6%.
Antonym vs Synonym Distinction using InterlaCed Encoder NETworks (ICE-NET) (2024.findings-eacl)

Copied to clipboard

Challenge: Existing research on antonym-synonym distinction is limited by the sparsity of the feature space.
Approach: They propose to capture and model relation-specific properties of antonyms and synonyms pairs . ICE-NET outperforms existing research by a relative score of upto 1.8% in F1-measure .
Outcome: The proposed model outperforms existing models by 1.8% in the F1-measure.
Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments (2025.emnlp-main)

Copied to clipboard

Challenge: Recent advances in Large Language Models (LLMs) have paved the way for VisionLarge Language Model (VLLM) capabilities have not been thoroughly explored in cluttered datasets where there is complex environment having deformedshaped objects.
Approach: They propose a dataset specifically designed for waste classification in real-world scenarios, characterized by complex environments and deformed shaped objects.
Outcome: The proposed dataset provides valuable insights into the performance of VLLMs under challenging conditions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations