Papers with Mozambique

2 papers
The PALMA Corpora of African Varieties of Portuguese (2022.lrec-1)

Copied to clipboard

Challenge: a corpus of urban varieties of Portuguese is being studied in Angola, Mozambique and So Tomé and Prncipe . the corpora are transcribed spoken data, complemented by metadata describing the setting of the audio recordings and sociolinguistic information about the speakers.
Approach: They present three new corpora of urban varieties of Portuguese spoken in Angola, Mozambique and So Tomé and Prncipe . they provide new, contemporary data for the study of each variety and for comparative research on African, Brazilian and European varieties .
Outcome: The corpora are transcribed spoken data and annotated with POS and lemma information . they are already being used for comparative research on possession and location .
Building Resources for Emakhuwa: Machine Translation and News Classification Benchmarks (2024.emnlp-main)

Copied to clipboard

Challenge: Emakhuwa is the most widely spoken language in Mozambique but has received limited attention in NLP research.
Approach: They propose a comprehensive collection of NLP resources for Emakhuwa, Mozambique's most widely spoken language.
Outcome: The proposed models show good performance in news topic classification and promising results in machine translation.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations