4 citations · 4 across the 5 of their papers we have counts for
5 papers
Data Center Audio/Video Intelligence on Device (DAVID) -- An Edge-AI Platform for Smart-Toys
Gabriel Cosache, Francisco Salgado, Cosmin Rotariu +3
An overview is given of the DAVID Smart-Toy platform, one of the first Edge AI platform designs to incorporate advanced low-power data processing by neural inference models co-loca…
Synthetic Speaking Children -- Why We Need Them and How to Make Them
Muhammad Ali Farooq, Dan Bigioi, Rishabh Jain +3
Contemporary Human Computer Interaction (HCI) research relies primarily on neural network models for machine vision and speech understanding of a system user. Such models require e…
A comparative analysis between Conformer-Transducer, Whisper, and wav2vec2 for improving the child speech recognition
Andrei Barcovschi, Rishabh Jain, Peter Corcoran
Automatic Speech Recognition (ASR) systems have progressed significantly in their performance on adult speech data; however, transcribing child speech remains challenging due to th…
Improved Child Text-to-Speech Synthesis through Fastpitch-based Transfer Learning
Rishabh Jain, Peter Corcoran
Speech synthesis technology has witnessed significant advancements in recent years, enabling the creation of natural and expressive synthetic speech. One area of particular interes…
Adaptation of Whisper models to child speech recognition
Rishabh Jain, Andrei Barcovschi, Mariam Yiwere +2
Automatic Speech Recognition (ASR) systems often struggle with transcribing child speech due to the lack of large child speech datasets required to accurately train child-friendly…