5 papers · 1 filter
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
Francesco Nespoli, Daniel Barreda, Patrick A. Naylor
In recent years, automatic speech recognition (ASR) models greatly improved transcription performance both in clean, low noise, acoustic conditions and in reverberant environments.…
Long-Term Conversation Analysis: Privacy-Utility Trade-off under Noise and Reverberation
Jule Pohlhausen, Francesco Nespoli, Joerg Bitzer
Recordings in everyday life require privacy preservation of the speech content and speaker identity. This contribution explores the influence of noise and reverberation on the trad…
Long-term Conversation Analysis: Exploring Utility and Privacy
Francesco Nespoli, Jule Pohlhausen, Patrick A. Naylor +1
The analysis of conversations recorded in everyday life requires privacy protection. In this contribution, we explore a privacy-preserving feature extraction method based on input…
Two-Stage Voice Anonymization for Enhanced Privacy
Francesco Nespoli, Daniel Barreda, Joerg Bitzer +1
In recent years, the need for privacy preservation when manipulating or storing personal data, including speech , has become a major issue. In this paper, we present a system addre…
Relative Acoustic Features for Distance Estimation in Smart-Homes
Francesco Nespoli, Daniel Barreda, Patrick A. Naylor
Any audio recording encapsulates the unique fingerprint of the associated acoustic environment, namely the background noise and reverberation. Considering the scenario of a room eq…