Papers by Marco Tagliasacchi
MAD Speech: Measures of Acoustic Diversity of Speech (2025.naacl-long)
Copied to clipboard
| Challenge: | Recent advances in generative spoken language modeling have produced models that produce speech in a wide range of voices, prosody and recording conditions. |
| Approach: | They propose acoustic diversity metrics that measure voice, gender, emotion, accent, background noise and a priori known diversity preferences for each facet. |
| Outcome: | The proposed metrics show that they achieve stronger agreement with diversity than baselines. |