4 papers
A two-step approach to leverage contextual data: speech recognition in air-traffic communications
Iuliia Nigmatulina, Juan Zuluaga-Gomez, Amrutha Prasad +2
Automatic Speech Recognition (ASR), as the assistance of speech communication between pilots and air-traffic controllers, can significantly reduce the complexity of the task and in…
Speech Activity Detection Based on Multilingual Speech Recognition System
Seyyed Saeed Sarfjoo, Srikanth Madikeri, Petr Motlicek
To better model the contextual information and increase the generalization ability of Speech Activity Detection (SAD) system, this paper leverages a multi-lingual Automatic Speech…
Graph2Speak: Improving Speaker Identification using Network Knowledge in Criminal Conversational Data
Mael Fabien, Seyyed Saeed Sarfjoo, Petr Motlicek +1
Criminal investigations mostly rely on the collection of speech conversational data in order to identify speakers and build or enrich an existing criminal network. Social network a…
Transformation of low-quality device-recorded speech to high-quality speech using improved SEGAN model
Seyyed Saeed Sarfjoo, Xin Wang, Gustav Eje Henter +3
Nowadays vast amounts of speech data are recorded from low-quality recorder devices such as smartphones, tablets, laptops, and medium-quality microphones. The objective of this res…