Papers by Shatha Altammami

2 papers
Challenging the Transformer-based models with a Classical Arabic dataset: Quran and Hadith (2022.lrec-1)

Copied to clipboard

Challenge: Existing benchmark datasets have a low readability index which does not reflect real-world complex data.
Approach: They constructed a dataset of Quran-verse and Hadith-teaching pairs by consulting sources of reputable religious experts.
Outcome: The proposed models performed on a binary classification task to identify whether two pieces of CA text convey the same underlying message.
Constructing a Bilingual Hadith Corpus Using a Segmentation Tool (2020.lrec-1)

Copied to clipboard

Challenge: Existing studies on Hadith have focused on the Quran, leaving it relatively unexplored.
Approach: They propose to gather and construct a bilingual parallel corpus of Islamic Hadith using a custom segmentation tool that annotates the two Hadithe components with 92% accuracy.
Outcome: The proposed method minimises the costs of language resource creation and produces consistent results independently from previous knowledge and experiences that usually influence human annotators.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations