3 citations · 5 across the 3 of their papers we have counts for
5 papers · 1 filter
Audio Retrieval with WavText5K and CLAP Training
Soham Deshmukh, Benjamin Elizalde, Huaming Wang
Audio-Text retrieval takes a natural language query to retrieve relevant audio files in a database. Conversely, Text-Audio retrieval takes an audio file as a query to retrieve rele…
Improving weakly supervised sound event detection with self-supervised auxiliary tasks
Soham Deshmukh, Bhiksha Raj, Rita Singh
While multitask and transfer learning has shown to improve the performance of neural networks in limited data settings, they require pretraining of the model on large datasets befo…
Interpreting glottal flow dynamics for detecting COVID-19 from voice
Soham Deshmukh, Mahmoud Al Ismail, Rita Singh
In the pathogenesis of COVID-19, impairment of respiratory functions is often one of the key symptoms. Studies show that in these cases, voice production is also adversely affected…
Detection of COVID-19 through the analysis of vocal fold oscillations
Mahmoud Al Ismail, Soham Deshmukh, Rita Singh
Phonation, or the vibration of the vocal folds, is the primary source of vocalization in the production of voiced sounds by humans. It is a complex bio-mechanical process that is h…
Multi-Task Learning for Interpretable Weakly Labelled Sound Event Detection
Soham Deshmukh, Bhiksha Raj, Rita Singh
Weakly Labelled learning has garnered lot of attention in recent years due to its potential to scale Sound Event Detection (SED) and is formulated as Multiple Instance Learning (MI…