| Challenge: | Aerodynamic processes underlie the characteristics of the acoustic signal of speech sounds. |
| Approach: | a database of aerodynamic processes underlies the characteristics of the acoustic signal of speech sounds . a project was undertaken to obtain data with simultaneous recording of speech acustic signals . |
| Outcome: | the database was designed during an ARC project . it contains recordings of 2 English, 1 Amharic, and 7 French speakers . |
Similar Papers
PRODIS - a Speech Database and a Phoneme-based Language Model for the Study of Predictability Effects in Polish (2024.lrec-main)
Copied to clipboard
| Challenge: | acoustic predictability is operationalised by surprisal in Polish, but cross-linguistic differences depend on prosodic system. |
| Approach: | They present a speech database and a phoneme-level language model of Polish . they aim to study contextual predictability effects on acoustic distinctiveness . |
| Outcome: | The proposed model is the first large, publicly available speech database of Polish . it is based on a light GPT architecture and can be expanded to other languages . |
The Speed-Vel Project: a Corpus of Acoustic and Aerodynamic Data to Measure Droplets Emission During Speech Interaction (2022.lrec-1)
Copied to clipboard
Francesca Carbone, Gilles Bouchet, Alain Ghio, Thierry Legou, Carine André, Muriel Lalain, Sabrina Kadri, Caterina Petrone, Federica Procino, Antoine Giovanni
| Challenge: | Conversations and professional interactions are associated with increased risk of SARS-CoV-2 exposure . however, it is unclear to what extent speech properties influence droplets emission . |
| Approach: | They propose to measure velocity and direction of airflow, the number and size of droplets spread during conversation in french. |
| Outcome: | The results will allow future simulation studies to predict the transport, dispersion and evaporation of droplets emitted under different speech conditions. |
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology (2026.findings-acl)
Copied to clipboard
| Challenge: | Dialectal Arabic datasets embody a range of domain, dialect, and quality. |
| Approach: | They propose a framework for automatic speech recognition in dialectal Arabic to address the limited data availability encountered in dialects. |
| Outcome: | The proposed framework provides access to 31 datasets covering 14 dialects to better address the limited data availability encountered in dialectal Arabic speech processing. |
Open-source Multi-speaker Corpora of the English Accents in the British Isles (2020.lrec-1)
Copied to clipboard
| Challenge: | Using a dataset of high-quality audio, the authors examine the accents of 120 volunteers in the British Isles. |
| Approach: | They present a dataset of high-quality audio of English sentences recorded by volunteers with different accents of the British Isles. |
| Outcome: | The transcribed audio includes pronunciations of global locations, major airlines and common personal names in different accents. |
The MonPaGe_HA Database for the Documentation of Spoken French Throughout Adulthood (L18-1)
Copied to clipboard
| Challenge: | Existing studies on life-span changes in the speech of adults are mainly based on English speakers and few studies have compared more than two extreme age groups. |
| Approach: | They describe a MonPaGe_HealthyAdults database of spoken french with 405 speakers aged from 20 to 93 years old. |
| Outcome: | The proposed database includes 405 speakers aged 20 to 93 years old and includes 4 regiolects. |
PATATRA and PATAFreq: two French databases for the documentation of within-speaker variability in speech (2022.lrec-1)
Copied to clipboard
| Challenge: | Variability in speech is pervasive but structured and ruled-governed. |
| Approach: | They propose two databases which contain recordings of 9 to 11 speakers . they compare the delay between repetitions of speech tasks with different speakers based on their own data . |
| Outcome: | The proposed databases compare speakers' performance on a large set of speech tasks with different delays. |
Design and Development of Speech Corpora for Air Traffic Control Training (L18-1)
Copied to clipboard
| Challenge: | The current state-of-the-art training procedures involve retired pilots that train as virtual plane pilots and process the spoken prompts to form that can be entered into software that simulates the plane movement on the radar screen. |
| Approach: | They describe the process of creating domain-specific speech corpora containing air traffic control (ATC) communication prompts. |
| Outcome: | The proposed system could be used for training air traffic controllers in the Czech Republic. |
A Manually Annotated Resource for the Investigation of Nasal Grunts (2020.lrec-1)
Copied to clipboard
| Challenge: | acoustic annotation of nasal grunts is described in the whole CID corpus of the french language . acculturation of non-lexical conversational sounds has been debated for a long time . |
| Approach: | They propose an annotation framework for nasal grunts of the whole French CID corpus . they characterise acoustic cues and visual cue conventions followed for the annotation . |
| Outcome: | The proposed framework is based on the entire French CID corpus. |
Corpus Creation and Automatic Alignment of Historical Dutch Dialect Speech (2024.lrec-main)
Copied to clipboard
| Challenge: | The Dutch Dialect Database contains dialectal variations of Dutch recorded in the second half of the twentieth century. |
| Approach: | They propose to create a corpus containing audio recordings and orthographic transcriptions of Dutch dialects recorded in the second half of the 20th century. |
| Outcome: | The Dutch Dialect Database contains dialectal variations recorded all over the Netherlands in the second half of the twentieth century. |
Praaline: An Open-Source System for Managing, Annotating, Visualising and Analysing Speech Corpora (P18-4)
Copied to clipboard
| Challenge: | Praaline is an open-source software system for constituting and managing spoken language and multimodal corpora. |
| Approach: | They present the latest developments of Praaline, an open-source software system for constituting and managing spoken language and multimodal corpora. |
| Outcome: | The proposed system can be used for creating, managing, visualising and analysing spoken language and multimodal corpora. |