2 citations · 5 across the 8 of their papers we have counts for
3 papers · 1 filter
Whispering LLaMA: A Cross-Modal Generative Error Correction Framework for Speech Recognition
Srijith Radhakrishnan, Chao-Han Huck Yang, Sumeer Ahmad Khan +4
We introduce a new cross-modal fusion technique designed for generative error correction in automatic speech recognition (ASR). Our methodology leverages both acoustic information…
A Parameter-Efficient Learning Approach to Arabic Dialect Identification with Pre-Trained General-Purpose Speech Model
Srijith Radhakrishnan, Chao-Han Huck Yang, Sumeer Ahmad Khan +3
In this work, we explore Parameter-Efficient-Learning (PEL) techniques to repurpose a General-Purpose-Speech (GSM) model for Arabic dialect identification (ADI). Specifically, we i…
IHCV: Discovery of Hidden Time-Dependent Control Variables in Non-Linear Dynamical Systems
Juan Munoz, Subash Balsamy, Juan P. Bernal-Tamayo +6
Discovering non-linear dynamical models from data is at the core of science. Recent progress hinges upon sparse regression of observables using extensive libraries of candidate fun…