4 papers
State-Space Models in Efficient Whispered and Multi-dialect Speech Recognition
Aref Farhadipour, Homayoon Beigi, Volker Dellwo +1
Whispered speech recognition presents significant challenges for conventional automatic speech recognition systems, particularly when combined with dialect variation. However, util…
MLP, XGBoost, KAN, TDNN, and LSTM-GRU Hybrid RNN with Attention for SPX and NDX European Call Option Pricing
Boris Ter-Avanesov, Homayoon Beigi
We explore the performance of various artificial neural network architectures, including a multilayer perceptron (MLP), Kolmogorov-Arnold network (KAN), LSTM-GRU hybrid recursive n…
Spontaneous Informal Speech Dataset for Punctuation Restoration
Xing Yi Liu, Homayoon Beigi
Presently, punctuation restoration models are evaluated almost solely on well-structured, scripted corpora. On the other hand, real-world ASR systems and post-processing pipelines…
Carnatic Raga Identification System using Rigorous Time-Delay Neural Network
Sanjay Natesan, Homayoon Beigi
Large scale machine learning-based Raga identification continues to be a nontrivial issue in the computational aspects behind Carnatic music. Each raga consists of many unique and…