Papers by Marie Bauer
Automatically Estimating Textual and Phonemic Complexity for Cued Speech: How to See the Sounds from French Texts (2024.lrec-main)
Copied to clipboard
| Challenge: | Cued Speech (CS) is a visual communication system developed for people with hearing loss to complement speech reading at the phonetic level. |
| Approach: | They propose a method to phonemize written corpora so that each word is aligned with the corresponding CS key(s) this method is part of a wider project aimed at creating an augmented reality system displaying a virtual coding hand where the user will be able to choose a text upon its complexity for cueing. |
| Outcome: | The proposed method is part of a wider project aimed at creating an augmented reality system displaying a virtual coding hand where the user can choose a text upon its complexity for cueing. |
Comprehensive Study on German Language Models for Clinical and Biomedical Text Understanding (2024.lrec-main)
Copied to clipboard
Ahmad Idrissi-Yaghir, Amin Dada, Henning Schäfer, Kamyar Arzideh, Giulia Baldini, Jan Trienes, Max Hasin, Jeanette Bewersdorff, Cynthia S. Schmidt, Marie Bauer, Kaleb E. Smith, Jiang Bian, Yonghui Wu, Jörg Schlötterer, Torsten Zesch, Peter A. Horn, Christin Seifert, Felix Nensa, Jens Kleesiek, Christoph M. Friedrich
| Challenge: | Pre-trained language models can struggle in specialized domains such as medicine . existing generalpurpose pre-tried models can be used and refined through further pre-training on domainspecific unlabeled data. |
| Approach: | They pre-trained German medical language models on 2.4B tokens from translated public data and 3B token of German clinical data. |
| Outcome: | The proposed models outperform clinical models on various downstream tasks in germany . the authors show that continuous pre-training can match or exceed clinical models trained from scratch . |