6 citations · 9 across the 18 of their papers we have counts for
10 papers · 1 filter
XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection
Yang Xiao, Rohan Kumar Das
Transformers and their variants have achieved great success in speech processing. However, their multi-head self-attention mechanism is computationally expensive. Therefore, one no…
Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection
Han Yin, Yang Xiao, Jisheng Bai +1
Sound Event Detection (SED) is challenging in noisy environments where overlapping sounds obscure target events. Language-queried audio source separation (LASS) aims to isolate the…
Exploring Text-Queried Sound Event Detection with Audio Source Separation
Han Yin, Jisheng Bai, Yang Xiao +6
In sound event detection (SED), overlapping sound events pose a significant challenge, as certain events can be easily masked by background noise or other events, resulting in poor…
TF-Mamba: A Time-Frequency Network for Sound Source Localization
Yang Xiao, Rohan Kumar Das
Sound source localization (SSL) determines the position of sound sources using multi-channel audio data. It is commonly used to improve speech enhancement and separation. Extractin…
Where's That Voice Coming? Continual Learning for Sound Source Localization
Yang Xiao, Rohan Kumar Das
Sound source localization (SSL) is essential for many speech-processing applications. Deep learning models have achieved high performance, but often fail when the training and infe…
UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection
Yang Xiao, Rohan Kumar Das
This work explores class-incremental learning (CIL) for sound event detection (SED), advancing adaptability towards real-world scenarios. CIL's success in domains like computer vis…