ATC-ANNO: Semantic Annotation for Air Traffic Control with Assistive Auto-Annotation (2020.lrec-1)
Copied to clipboard
| Challenge: | ATC communications are a challenging domain for automatic speech recognition (ASR) due to the time-sensitive nature of their task, annotators must have prior experience with ATC communication. |
| Approach: | They propose a tool for the transcription and semantic annotation of air traffic communications. |
| Outcome: | The proposed tool can annotate four times as many utterances in a single time. |
Similar Papers
Design and Development of Speech Corpora for Air Traffic Control Training (L18-1)
Copied to clipboard
| Challenge: | The current state-of-the-art training procedures involve retired pilots that train as virtual plane pilots and process the spoken prompts to form that can be entered into software that simulates the plane movement on the radar screen. |
| Approach: | They describe the process of creating domain-specific speech corpora containing air traffic control (ATC) communication prompts. |
| Outcome: | The proposed system could be used for training air traffic controllers in the Czech Republic. |
A Real-life, French-accented Corpus of Air Traffic Control Communications (L18-1)
Copied to clipboard
Estelle Delpech, Marion Laignelet, Christophe Pimm, Céline Raynal, Michal Trzos, Alexandre Arnold, Dominique Pronto
| Challenge: | AIRBUS-ATC corpus is a real-life, french-accented speech corpus of air traffic control (ATC) communications . it is composed of 59 hours of transcribed English audio, along with linguistic and meta-data annotations. |
| Approach: | They propose to use a real-life, French-accented speech corpus of ATC communications to build a robust ATC speech recognition engine. |
| Outcome: | The AIRBUS-ATC corpus is composed of 59 hours of transcribed English audio, along with linguistic and meta-data annotations. |
Knowledge extraction from aeronautical messages (NOTAMs) with self-supervised language models for aircraft pilots (2022.naacl-industry)
Copied to clipboard
| Challenge: | During pre-flight briefings, aircraft pilots analyse a long list of NOTAMs . the messages are usually written in the English language, but the phrasing is very special . |
| Approach: | They pretrain language models derived from BERT on circa 1 million unlabeled NOTAMs . they reuse the learnt representations on three downstream tasks valuable for pilots - criticality prediction, named entity recognition and translation into a structured language called Airlang. |
| Outcome: | The proposed language model can be used on criticality prediction, named entity recognition and translation into a structured language called Airlang. |
ALANNO: An Active Learning Annotation System for Mortals (2023.eacl-demo)
Copied to clipboard
| Challenge: | Active learning (AL) is a special family of machine learning algorithms designed to reduce labeling costs and improve accuracy. |
| Approach: | They developed an open-source annotation system for NLP tasks equipped with features to make AL effective in real-world annotation projects. |
| Outcome: | ALANNO is an open-source annotation system for NLP tasks equipped with features to make AL effective in real-world annotation projects. |
ATGen: A Framework for Active Text Generation (2025.acl-demo)
Copied to clipboard
Akim Tsvigun, Daniil Vasilev, Ivan Tsvigun, Ivan Lysenko, Talgat Bektleuov, Aleksandr Medvedev, Uliana Vinogradova, Nikita Severin, Mikhail Mozikov, Andrey Savchenko, Ilya Makarov, Grigorev Rostislav, Ramil Kuleev, Fedor Zhdanov, Artem Shelmanov
| Challenge: | Despite the surging popularity of natural language generation tasks, the application of active learning (AL) to NLG has been limited. |
| Approach: | They propose a framework that bridges AL with text generation tasks and provides a unified platform for smooth implementation and benchmarking of novel AL strategies tailored to NLG tasks. |
| Outcome: | The proposed framework simplifies AL-empowered annotation in NLG tasks using both human annotators and automatic annotation agents based on large language models (LLMs). |
Praat++: Multimedia Annotation System for Speech and Vocalization (2026.acl-demo)
Copied to clipboard
| Challenge: | High-quality time-aligned annotation is fundamental to speech processing and animal vocalization research, yet precise boundary localization and consistent labeling remain challenging in collaborative settings. |
| Approach: | They propose a web-based multimedia annotation system for collaborative, video-informed, and AI-assisted timeline labeling of audio and video data. |
| Outcome: | The proposed system improves time-aligned labeling and accuracy in speech and animal vocalization annotations. |
Universal Semantic Annotator: the First Unified API for WSD, SRL and Semantic Parsing (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing approaches to understanding textual information are still far from achieving true natural language understanding (NLU). |
| Approach: | They propose a unified API for high-quality automatic annotations of texts in 100 languages through state-of-the-art systems for Word Sense Disambiguation, Semantic Role Labeling and Semantics Parsing. |
| Outcome: | The proposed system can provide users with rich and diverse semantic information, help second-language learners, and integrate explicit semantic knowledge into downstream tasks and real-world applications. |
X-ACE: Explainable and Multi-factor Audio Captioning Evaluation (2024.findings-acl)
Copied to clipboard
| Challenge: | Existing evaluation metrics for automated audio captioning only provide an overall score . current evaluation checklists are inadequate to characterize the nuanced differences . |
| Approach: | They propose an explainable and multi-factor audio captioning evaluation paradigm . they define sound event, source, attribute and relation as four factors tailored for the audio description . |
| Outcome: | The proposed evaluation paradigm improves the quality of audio captions . it can detect mismatches and align with human perception, the authors show . |
AnnoTheia: A Semi-Automatic Annotation Toolkit for Audio-Visual Speech Technologies (2024.lrec-main)
Copied to clipboard
| Challenge: | a small fraction of the languages currently covered by speech technologies are mainly spoken in English. |
| Approach: | They present an annotation toolkit that detects when a person speaks on the scene and the corresponding transcription. |
| Outcome: | The proposed toolkit can speed up the annotation process by up to four times . it can be used in Spanish, and is available on github. |
Rationally Reappraising ATIS-based Dialogue Systems (P19-1)
Copied to clipboard
| Challenge: | Recent state-of-the-art neural models have obtained F1-scores near 98% on the task of slot filling. |
| Approach: | They propose to fix annotation errors in ATIS and propose a rule-based grammar for slot filling that achieves a 95.82% F1 score. |
| Outcome: | The proposed grammar achieves a 95.82% F1-score on the ATIS domain. |