1 citations · 1 across the 4 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2024★ 1 cited
LipGER: Visually-Conditioned Generative Error Correction for Robust Automatic Speech Recognition
Sreyan Ghosh, Sonal Kumar, Ashish Seth +4
Visual cues, like lip motion, have been shown to improve the performance of Automatic Speech Recognition (ASR) systems in noisy environments. We propose LipGER (Lip Motion aided Ge…
eess.AS2023
UNFUSED: UNsupervised Finetuning Using SElf supervised Distillation
Ashish Seth, Sreyan Ghosh, S. Umesh +1
In this paper, we introduce UnFuSeD, a novel approach to leverage self-supervised learning and reduce the need for large amounts of labeled data for audio classification. Unlike pr…