Papers by Hafte Abera
Large Vocabulary Read Speech Corpora for Four Ethiopian Languages: Amharic, Tigrigna, Oromo and Wolaytta (2020.lrec-1)
Copied to clipboard
Solomon Teferra Abate, Martha Yifiru Tachbelie, Michael Melese, Hafte Abera, Tewodros Abebe, Wondwossen Mulugeta, Yaregal Assabie, Million Meshesha, Solomon Afnafu, Binyam Ephrem Seyoum
| Challenge: | Automatic Speech Recognition (ASR) is one of the most important technologies to support spoken communication in modern life. |
| Approach: | They have developed four large speech corpora for four Ethiopian languages . they have word error rates of 37.65%, 31.03%, 38.02%, 33.89% for each language . |
| Outcome: | The proposed corpora achieve word error rates of 37.65%, 31.03%, 38.02%, 33.89% for Amharic, Tigrigna, Oromo and Wolaytta. |
Parallel Corpora for bi-lingual English-Ethiopian Languages Statistical Machine Translation (C18-1)
Copied to clipboard
Solomon Teferra Abate, Michael Melese, Martha Yifiru Tachbelie, Million Meshesha, Solomon Atinafu, Wondwossen Mulugeta, Yaregal Assabie, Hafte Abera, Binyam Ephrem, Tewodros Abebe, Wondimagegnhue Tsegaye, Amanuel Lemma, Tsegaye Andargie, Seifedin Shifaw
| Challenge: | Various approaches to machine translation have been and are being used in the research community, that can broadly classified as rule-based and corpus based. |
| Approach: | They propose to develop parallel corpora for English and Ethiopian languages such as Amharic, Tigrigna, Afan-Oromo, Wolaytta and Ge’ez. |
| Outcome: | The proposed system improves on the English-Ethiopian languages. |