6 papers
A Feature Engineering Approach for Literary and Colloquial Tamil Speech Classification using 1D-CNN
M. Nanmalar, S. Johanan Joysingh, P. Vijayalakshmi +1
In ideal human computer interaction (HCI), the colloquial form of a language would be preferred by most users, since it is the form used in their day-to-day conversations. However,…
Development of Large Annotated Music Datasets using HMM-based Forced Viterbi Alignment
S. Johanan Joysingh, P. Vijayalakshmi, T. Nagarajan
Datasets are essential for any machine learning task. Automatic Music Transcription (AMT) is one such task, where considerable amount of data is required depending on the way the s…
MaskCycleGAN-based Whisper to Normal Speech Conversion
K. Rohith Gupta, K. Ramnath, S. Johanan Joysingh +2
Whisper to normal speech conversion is an active area of research. Various architectures based on generative adversarial networks have been proposed in the recent past. Especially,…
Quartered Chirp Spectral Envelope for Whispered vs Normal Speech Classification
S. Johanan Joysingh, P. Vijayalakshmi, T. Nagarajan
Whispered speech as an acceptable form of human-computer interaction is gaining traction. Systems that address multiple modes of speech require a robust front-end speech classifier…
Quartered Spectral Envelope and 1D-CNN-based Classification of Normally Phonated and Whispered Speech
S. Johanan Joysingh, P. Vijayalakshmi, T. Nagarajan
Whisper, as a form of speech, is not sufficiently addressed by mainstream speech applications. This is due to the fact that systems built for normal speech do not work as expected…
Chirp Group Delay based Onset Detection in Instruments with Fast Attack
S. Johanan Joysingh, P. Vijayalakshmi, T. Nagarajan
The onset of a musical note is the earliest time at which a note can be reliably detected. Detection of these musical onsets pose challenges in the presence of ornamentation such a…