activity
20112022
most citedOptimization of data-driven filterbank for automatic speaker verification

45 citations · 120 across the 9 of their papers we have counts for

collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS202211 cited

Analysis of constant-Q filterbank based representations for speech emotion recognition

Premjeet Singh, Shefali Waldekar, Md Sahidullah +1

This work analyzes the constant-Q filterbank-based time-frequency representations for speech emotion recognition (SER). Constant-Q filterbank provides non-linear spectro-temporal r…

eess.AS2021

Cross-Corpora Language Recognition: A Preliminary Investigation with Indian Languages

Spandan Dey, Goutam Saha, Md Sahidullah

In this paper, we conduct one of the very first studies for cross-corpora performance evaluation in the spoken language identification (LID) problem. Cross-corpora evaluation was n…

eess.AS2021

Deep scattering network for speech emotion recognition

Premjeet Singh, Goutam Saha, Md Sahidullah

This paper introduces scattering transform for speech emotion recognition (SER). Scattering transform generates feature representations which remain stable to deformations and shif…

eess.AS2021

Non-linear frequency warping using constant-Q transformation for speech emotion recognition

Premjeet Singh, Goutam Saha, Md Sahidullah

In this work, we explore the constant-Q transform (CQT) for speech emotion recognition (SER). The CQT-based time-frequency analysis provides variable spectro-temporal resolution wi…

eess.AS202045 cited

Optimization of data-driven filterbank for automatic speaker verification

Susanta Sarangi, Md Sahidullah, Goutam Saha

Most of the speech processing applications use triangular filters spaced in mel-scale for feature extraction. In this paper, we propose a new data-driven filter design method which…