Charles Translator: A Machine Translation System between Ukrainian and Czech (2024.lrec-main)
Copied to clipboard
Martin Popel, Lucie Polakova, Michal Novák, Jindřich Helcl, Jindřich Libovický, Pavel Straňák, Tomas Krabac, Jaroslava Hlavacova, Mariia Anisimova, Tereza Chlanova
| Challenge: | a system for translating between Ukrainian and Czech was developed in the spring of 2022 . the system was not available at the time in the required quality . |
| Approach: | They propose a machine translation system between Ukrainian and Czech to reduce the impact of the Russian-Ukrainian war on individuals and society. |
| Outcome: | The proposed system translates directly between Ukrainian and Czech, compared to other systems that use English as a pivot. |
Similar Papers
Long to reign over us: A Case Study of Machine Translation and a New Monarch (2023.findings-acl)
Copied to clipboard
| Challenge: | We examine translations between French and English in contexts with ambiguity . with the passing of Queen Elizabeth II, MT systems can produce errors due to linguistic features of both languages and the paucity of references to kings in the training data. |
| Approach: | They examine translations between French and English as they were produced by MT systems . they find that even when human translators would have adequate context, machine translation systems do not always produce the expected output. |
| Outcome: | The proposed model shows that even when human translators have context, machine translation systems do not always produce the expected output. |
Demonstration of a Neural Machine Translation System with Online Learning for Translators (P19-3)
Copied to clipboard
Miguel Domingo, Mercedes García-Martínez, Amando Estela Pastor, Laurent Bié, Alexander Helle, Álvaro Peris, Francisco Casacuberta, Manuel Herranz Pérez
| Challenge: | a new method of "humanizing" automatic translations has been developed for the translation industry . a demonstration of an online learning system for machine translation in a production environment . |
| Approach: | They present a system which implements online learning for neural machine translation in a production environment. |
| Outcome: | The proposed system saves post-editing effort and adapts to a specific domain or user style. |
Dialectal and Low Resource Machine Translation for Aromanian (2025.coling-main)
Copied to clipboard
| Challenge: | Existing training methods for low-resource languages are focused on English or are massively multilingual, but do not consider the particularities of lowresource language. |
| Approach: | They propose a neural machine translation system that can translate between Romanian, English, and Aromanian. |
| Outcome: | The proposed system can translate between Romanian, English, and Aromanian . BLEU scores range from 17 to 32 depending on direction and genre of text . |
Many-to-English Machine Translation Tools, Data, and Pretrained Models (2021.acl-demo)
Copied to clipboard
| Challenge: | Commercial translation systems support only one hundred languages or fewer . commercial translation systems do not make these models available for transfer to low resource languages . |
| Approach: | They propose a multilingual neural machine translation model that can translate from 500 source languages to English. |
| Outcome: | The proposed model can translate from 500 source languages to English, or be used as a parent model for low-resource languages. |
Multilingual Neural Machine Translation (2020.coling-tutorials)
Copied to clipboard
| Challenge: | In this tutorial, we will cover the latest advances in NMT to enhance low-resource translation. |
| Approach: | They will cover the latest advances in NMT approaches that leverage multilingualism . they will focus on topics such as language divergence, transfer learning and pivoting . |
| Outcome: | This tutorial will cover the latest advances in NMT to enhance low-resource translation models. |
A Tulu Resource for Machine Translation (2024.lrec-main)
Copied to clipboard
| Challenge: | Using parallel datasets, we train a machine translation system in English–Tulu . |
| Approach: | They present a parallel dataset for English–Tulu translation using human translations into the multilingual machine translation resource FLORES-200. |
| Outcome: | The proposed model outperforms Google Translate by 19 BLEU points (in September 2023). |
Data Augmentation Techniques for Machine Translation of Code-Switched Texts: A Comparative Study (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Code-switching (CSW) text generation is a popular solution to address data scarcity. |
| Approach: | They compare linguistic theories, lexical replacements and back-translation approaches to Egyptian Arabic-English CSW. |
| Outcome: | The proposed methods perform best on machine translation and quality evaluation. |
Selecting Machine-Translated Data for Quick Bootstrapping of a Natural Language Understanding System (N18-3)
Copied to clipboard
| Challenge: | In recent years, there has been growing interest in voice-controlled devices, such as Amazon Alexa or Google home. |
| Approach: | They investigate the use of Machine Translation to bootstrap a natural language understanding system for a new language for the use case of a large-scale voice-controlled device. |
| Outcome: | The proposed method reduces the time and cost of getting annotated corpus for a new language while still providing a large enough coverage of user requests. |
English-Basque Statistical and Neural Machine Translation (L18-1)
Copied to clipboard
| Challenge: | Neural machine translation (NMT) requires large training corpora, which is problematic for low-resource languages. |
| Approach: | They propose to use an open-domain and an IT-domain corpora to train machine translations in English-Basque. |
| Outcome: | The proposed systems outperform OpenNMT, Moses SMT and Google Translate in English-Basque translation. |
Selecting Backtranslated Data from Multiple Sources for Improved Neural Machine Translation (2020.acl-main)
Copied to clipboard
| Challenge: | incorporating backtranslated data from different sources has led to improved results in machine translation (MT) |
| Approach: | They use a low-resource use-case and a high-resourced language pair to test different backtranslation scenarios and employ data selection to optimise the synthetic corpora. |
| Outcome: | The proposed method reduces the amount of data used while maintaining high-quality MT systems. |