Papers by Ingmar Steiner
Creating New Language and Voice Components for the Updated MaryTTS Text-to-Speech Synthesis Platform (L18-1)
Copied to clipboard
| Challenge: | a reboot of the MaryTTS system became unavoidable due to the number of people who have contributed to its development over the years. |
| Approach: | They propose a workflow to create components for the MaryTTS text-to-speech synthesis platform. |
| Outcome: | The proposed workflow is compatible with the updated MaryTTS architecture, enabling new features and state-of-the-art paradigms such as synthesis based on deep neural networks (DNNs). |
A Multimodal Corpus of Expert Gaze and Behavior during Phonetic Segmentation Tasks (L18-1)
Copied to clipboard
| Challenge: | Phonetic segmentation is the process of splitting speech into distinct phonetic units . methods for automatic segmentation are not always accurate enough . |
| Approach: | They propose to model phonetic segmentation as close as possible to manual segmentation by recording experts performing a segmentation task. |
| Outcome: | This corpus captures human segmentation behavior by recording experts performing a segmentation task. |