2 citations · 2 across the 2 of their papers we have counts for
4 papers
Investigation of Whisper ASR Hallucinations Induced by Non-Speech Audio
Mateusz Barański, Jan Jasiński, Julitta Bartolewska +3
Hallucinations of deep neural models are amongst key challenges in automatic speech recognition (ASR). In this paper, we investigate hallucinations of the Whisper ASR model induced…
HeightCeleb - an enrichment of VoxCeleb dataset with speaker height information
Stanisław Kacprzak, Konrad Kowalczyk
Prediction of speaker's height is of interest for voice forensics, surveillance, and automatic speaker profiling. Until now, TIMIT has been the most popular dataset for training an…
Refining DNN-based Mask Estimation using CGMM-based EM Algorithm for Multi-channel Noise Reduction
Julitta Bartolewska, Stanisław Kacprzak, Konrad Kowalczyk
In this paper, we present a method that allows to further improve speech enhancement obtained with recently introduced Deep Neural Network (DNN) models. We propose a multi-channel…
Causal Signal-Based DCCRN with Overlapped-Frame Prediction for Online Speech Enhancement
Julitta Bartolewska, Stanisław Kacprzak, Konrad Kowalczyk
The aim of speech enhancement is to improve speech signal quality and intelligibility from a noisy microphone signal. In many applications, it is crucial to enable processing with…