Papers by Martin Popel
Do UD Trees Match Mention Spans in Coreference Annotations? (2021.findings-emnlp)
Copied to clipboard
| Challenge: | Existing methods to annotate mention spans are based on delimiting token intervals, but there is no syntactic representation of the mention span. |
| Approach: | They propose to integrate coreference annotation with syntactic annotation to make them convergent in the long term. |
| Outcome: | The proposed approach could be advantageous in the long term, the authors argue. |
Neural Machine Translation Quality and Post-Editing Performance (2021.emnlp-main)
Copied to clipboard
| Challenge: | a recent study has shown that MT post-editing can reduce translation quality and speed . a large-scale study involving 30 professional translators examined the relationship between MT performance and post-edited outputs. |
| Approach: | They examine the relationship between MT performance and post-editing time and quality . they use neural MT of high quality to improve translation quality based on phrase-based MT . |
| Outcome: | The proposed model is not stable predictor of time or quality, the authors say . they find that better MT systems lead to fewer changes in the sentences . |
Universal Anaphora: The First Three Years (2024.lrec-main)
Copied to clipboard
Massimo Poesio, Maciej Ogrodniczuk, Vincent Ng, Sameer Pradhan, Juntao Yu, Nafise Sadat Moosavi, Silviu Paun, Amir Zeldes, Anna Nedoluzhko, Michal Novák, Martin Popel, Zdeněk Žabokrtský, Daniel Zeman
| Challenge: | Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by expanding the aspects of anaphonic interpretation which are or can be reliably annotated in an anagraphic corpora. |
| Approach: | They propose to develop a standard for anaphoric annotations and a method for evaluating models that can carry out this type of interpretation. |
| Outcome: | The Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by producing unified standards to annotate and encode annotations, delivering datasets encoded according to these standards, and developing methods for evaluating models that carry out this type of interpretation. |
CorefUD 1.0: Coreference Meets Universal Dependencies (2022.lrec-1)
Copied to clipboard
| Challenge: | Recent advances in standardization for annotated language resources have led to successful large scale efforts, such as the Universal Dependencies (UD) project for multilingual syntactically annotized data. |
| Approach: | They propose a multilingual collection of corpora and a standardized format for coreference resolution compatible with morphosyntactic annotations in the UD framework. |
| Outcome: | The proposed framework is compatible with morphosyntactic annotations and includes facilities for related tasks such as named entity recognition. |
Charles Translator: A Machine Translation System between Ukrainian and Czech (2024.lrec-main)
Copied to clipboard
Martin Popel, Lucie Polakova, Michal Novák, Jindřich Helcl, Jindřich Libovický, Pavel Straňák, Tomas Krabac, Jaroslava Hlavacova, Mariia Anisimova, Tereza Chlanova
| Challenge: | a system for translating between Ukrainian and Czech was developed in the spring of 2022 . the system was not available at the time in the required quality . |
| Approach: | They propose a machine translation system between Ukrainian and Czech to reduce the impact of the Russian-Ukrainian war on individuals and society. |
| Outcome: | The proposed system translates directly between Ukrainian and Czech, compared to other systems that use English as a pivot. |