Papers by Saméh Kchaou
Standardisation of Dialect Comments in Social Networks in View of Sentiment Analysis : Case of Tunisian Dialect (2022.lrec-1)
Copied to clipboard
| Challenge: | Using the internet, the spoken Arabic dialect language becomes informal languages written in social media . this linguistic situation inhibits mutual understanding and makes computational approaches difficult . we present a pipeline to standardize the written texts in social networks by translating them to MSA . |
| Approach: | They propose a pipeline to standardize Arabic written texts by translating them to MSA . they use a bert-based model to select Tunisian Dialect from MSA and other dialects . |
| Outcome: | The proposed pipeline achieves the best score for the standardization of written texts in social networks . the proposed pipeline includes the translated TD and the original text written in MSA . |
Text and Speech-based Tunisian Arabic Sub-Dialects Identification (2020.lrec-1)
Copied to clipboard
| Challenge: | Dialect IDentification is a difficult task when it is about the identification of dialects belonging to the same country. |
| Approach: | They present results on a dialect classification task covering four sub-dialects spoken in Tunisia using a spoken corpus of 1673 utterances. |
| Outcome: | The proposed system achieves an F-1 score of 93.75% while the F-1 is limited to 54.16% using text-based DID on the same test set. |