Papers by Martin Popel

5 papers
Do UD Trees Match Mention Spans in Coreference Annotations? (2021.findings-emnlp)

Copied to clipboard

Challenge: Existing methods to annotate mention spans are based on delimiting token intervals, but there is no syntactic representation of the mention span.
Approach: They propose to integrate coreference annotation with syntactic annotation to make them convergent in the long term.
Outcome: The proposed approach could be advantageous in the long term, the authors argue.
Neural Machine Translation Quality and Post-Editing Performance (2021.emnlp-main)

Copied to clipboard

Challenge: a recent study has shown that MT post-editing can reduce translation quality and speed . a large-scale study involving 30 professional translators examined the relationship between MT performance and post-edited outputs.
Approach: They examine the relationship between MT performance and post-editing time and quality . they use neural MT of high quality to improve translation quality based on phrase-based MT .
Outcome: The proposed model is not stable predictor of time or quality, the authors say . they find that better MT systems lead to fewer changes in the sentences .
Universal Anaphora: The First Three Years (2024.lrec-main)

Copied to clipboard

Challenge: Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by expanding the aspects of anaphonic interpretation which are or can be reliably annotated in an anagraphic corpora.
Approach: They propose to develop a standard for anaphoric annotations and a method for evaluating models that can carry out this type of interpretation.
Outcome: The Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by producing unified standards to annotate and encode annotations, delivering datasets encoded according to these standards, and developing methods for evaluating models that carry out this type of interpretation.
CorefUD 1.0: Coreference Meets Universal Dependencies (2022.lrec-1)

Copied to clipboard

Challenge: Recent advances in standardization for annotated language resources have led to successful large scale efforts, such as the Universal Dependencies (UD) project for multilingual syntactically annotized data.
Approach: They propose a multilingual collection of corpora and a standardized format for coreference resolution compatible with morphosyntactic annotations in the UD framework.
Outcome: The proposed framework is compatible with morphosyntactic annotations and includes facilities for related tasks such as named entity recognition.
Charles Translator: A Machine Translation System between Ukrainian and Czech (2024.lrec-main)

Copied to clipboard

Challenge: a system for translating between Ukrainian and Czech was developed in the spring of 2022 . the system was not available at the time in the required quality .
Approach: They propose a machine translation system between Ukrainian and Czech to reduce the impact of the Russian-Ukrainian war on individuals and society.
Outcome: The proposed system translates directly between Ukrainian and Czech, compared to other systems that use English as a pivot.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations