paper-with-me

Papers

Speech and Natural Language Processing Technologies for Pseudo-Pilot Simulator

2022-12-14 · Amrutha Prasad, Juan Zuluaga-Gomez, Petr Motlicek, Saeed Sarfjoo, Iuliia Nigmatulina, Karel Vesely

This paper describes a simple yet efficient repetition-based modular system for speeding up air-traffic controllers (ATCos) training. E.g., a human pilot is still required in EUROCONTROL's ESCAPE lite simulator (see https://www.eurocontrol.int/simulator/escape) during ATCo training. However, this need can be substituted by an automatic system that could act as a pilot. In this paper, we aim to develop and integrate a pseudo-pilot agent into the ATCo training pipeline by merging diverse artificial intelligence (AI) powered modules. The system understands the voice communications issued by the ATCo, and, in turn, it generates a spoken prompt that follows the pilot's phraseology to the initial communication. Our system mainly relies on open-source AI tools and air traffic control (ATC) databases, thus, proving its simplicity and ease of replicability. The overall pipeline is composed of the following: (1) a submodule that receives and pre-processes the input stream of raw audio, (2) an automatic speech recognition (ASR) system that transforms audio into a sequence of words; (3) a high-level ATC-related entity parser, which extracts relevant information from the communication, i.e., callsigns and commands, and finally, (4) a speech synthesizer submodule that generates responses based on the high-level ATC entities previously extracted. Overall, we show that this system could pave the way toward developing a real proof-of-concept pseudo-pilot system. Hence, speeding up the training of ATCos while drastically reducing its overall cost.

📄 PDF Abstract BibTeX arXiv:2212.07164

Code (0)

등록된 구현이 없습니다.

Tasks

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

Similar Papers 제목 키워드 기반

Kallaama: A Transcribed Speech Dataset about Agriculture in the Three Most Widely Spoken Languages in Senegal

2024-04-02 · Elodie Gauthier, Aminata Ndiaye, Abdoulaye Guissé

This work is part of the Kallaama project, whose objective is to produce and disseminate national languages corpora for speech technologies developments, in the field of agriculture. Except for Wolof, which benefits from…

Automatic Speech Recognitionspeech-recognitionSpeech Recognition

Speech Pseudonymisation Assessment Using Voice Similarity Matrices

2020-08-30 · Paul-Gauthier Noé, Jean-François Bonastre, Driss Matrouf, Natalia Tomashenko 외

The proliferation of speech technologies and rising privacy legislation calls for the development of privacy preservation solutions for speech applications. These are essential since speech signals convey a wealth of ric…

De-identificationVoice Similarity

Global Open Resources and Information for Language and Linguistic Analysis (GORILLA)

2016-05-01 · LREC 2016 5 · Damir Cavar, Malgorzata Cavar, Lwin Moe

The infrastructure Global Open Resources and Information for Language and Linguistic Analysis (GORILLA) was created as a resource that provides a bridge between disciplines such as documentary, theoretical, and corpus li…

Mamba in Speech: Towards an Alternative to Self-Attention

2024-05-21 · Xiangyu Zhang, Qiquan Zhang, Hexin Liu, Tianyi Xiao 외

Transformer and its derivatives have achieved success in diverse tasks across computer vision, natural language processing, and speech processing. To reduce the complexity of computations within the multi-head self-atten…

MambaSpeech Enhancementspeech-recognitionSpeech Recognition+1

SpeechBrain: A General-Purpose Speech Toolkit

2021-06-08 · Mirco Ravanelli, Titouan Parcollet, Peter Plantinga, Aku Rouhe 외

SpeechBrain is an open-source and all-in-one speech toolkit. It is designed to facilitate the research and development of neural speech processing technologies by being simple, flexible, user-friendly, and well-documente…

Language IdentificationSpoken Language Understanding