Papers by Marco Tagliasacchi

1 papers
MAD Speech: Measures of Acoustic Diversity of Speech (2025.naacl-long)

Copied to clipboard

Challenge: Recent advances in generative spoken language modeling have produced models that produce speech in a wide range of voices, prosody and recording conditions.
Approach: They propose acoustic diversity metrics that measure voice, gender, emotion, accent, background noise and a priori known diversity preferences for each facet.
Outcome: The proposed metrics show that they achieve stronger agreement with diversity than baselines.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations