Papers by Gustavo Paetzold
Text Simplification from Professionally Produced Corpora (L18-1)
Copied to clipboard
| Challenge: | Existing approaches to Text Simplification rely on the Wikipedia-Simple Wikipedia parallel corpus, which is used for many tasks. |
| Approach: | They propose to use the Newsela corpus to extract 550, 644 complex-simple sentence pairs from the corpus and introduce a lexical simplifier that uses the corpu to generate candidate simplifications. |
| Outcome: | The proposed model outperforms state-of-the-art approaches and generates candidate simplifications from the newsela corpus. |
SimPA: A Sentence-Level Simplification Corpus for the Public Administration Domain (L18-1)
Copied to clipboard
| Challenge: | lexical simplification is the task of reducing lexically and/or structural complexity of texts. |
| Approach: | They propose to collect manual simplifications for 1,100 original sentences using a sentence-level corpus from the Public Administration domain. |
| Outcome: | The proposed corpus contains 1,100 original sentences with manual simplifications collected through a two-stage process. |
Lexi: A tool for adaptive, personalized text simplification (C18-1)
Copied to clipboard
| Challenge: | Existing research on text simplification has aimed to develop generic solutions . instead, we need to develop customized simplification systems for individual users . |
| Approach: | They propose a framework for adaptive lexical simplification and introduce Lexi, a free open-source tool for personalized text simplification. |
| Outcome: | The proposed framework is based on a free open-source tool for adaptive, personalized text simplification. |