Papers by Jens Edlund
Augmented Prompt Selection for Evaluation of Spontaneous Speech Synthesis (2020.lrec-1)
Copied to clipboard
| Challenge: | Spontaneous speech is unscripted and created on the fly by the speaker, whereas read speech is pre-planned. |
| Approach: | They propose a tool that allows developers to select a varied, representative set of utterances from a spoken genre to be used for evaluation of TTS for a given domain. |
| Outcome: | The proposed tool can be used to evaluate TTS for a given domain using visualisation and tree-based algorithm. |
Revisiting Three Text-to-Speech Synthesis Experiments with a Web-Based Audience Response System (2024.lrec-main)
Copied to clipboard
| Challenge: | Audience Response System (ARS) evaluations are not well understood for text-to-speech synthesis (TTS) evaluation is a key weakness in the field and needs to adapt to be better-suited for this new generation of voices. |
| Approach: | They revisit three published TTS studies and perform an ARS-based evaluation on the stimuli used in each study. |
| Outcome: | The results show that Audience Response System (ARS) is highly useful for evaluating long and continuous stimuli. |
Bringing Order to Chaos: A Non-Sequential Approach for Browsing Large Sets of Found Audio Data (L18-1)
Copied to clipboard
| Challenge: | a new approach to search for sound in large archives is being developed . speech and speech technology researchers struggle to access large amounts of data . |
| Approach: | They propose a method for fast and efficient non-sequential browsing of sound in large archives that we know little about . they combine audio browsing through massively multi-object sound environments and an unsupervised dimensionality reduction algorithm to search for sound in public archives. |
| Outcome: | The proposed method is shown to combine well, resulting in rapid and interpretable observations. |