activity
20242026
most citedEnvSDD: Benchmarking Environmental Sound Deepfake Detection

6 citations · 9 across the 18 of their papers we have counts for

collaborators
Showing 2024Show all

10 papers · 1 filter

eess.AS2024

XLSR-Mamba: A Dual-Column Bidirectional State Space Model for Spoofing Attack Detection

Yang Xiao, Rohan Kumar Das

Transformers and their variants have achieved great success in speech processing. However, their multi-head self-attention mechanism is computationally expensive. Therefore, one no…

eess.AS2024

Leveraging LLM and Text-Queried Separation for Noise-Robust Sound Event Detection

Han Yin, Yang Xiao, Jisheng Bai +1

Sound Event Detection (SED) is challenging in noisy environments where overlapping sounds obscure target events. Language-queried audio source separation (LASS) aims to isolate the…

eess.AS2024

Exploring Text-Queried Sound Event Detection with Audio Source Separation

Han Yin, Jisheng Bai, Yang Xiao +6

In sound event detection (SED), overlapping sound events pose a significant challenge, as certain events can be easily masked by background noise or other events, resulting in poor…

eess.AS2024

TF-Mamba: A Time-Frequency Network for Sound Source Localization

Yang Xiao, Rohan Kumar Das

Sound source localization (SSL) determines the position of sound sources using multi-channel audio data. It is commonly used to improve speech enhancement and separation. Extractin…

eess.AS2024★ 2 cited

Where's That Voice Coming? Continual Learning for Sound Source Localization

Yang Xiao, Rohan Kumar Das

Sound source localization (SSL) is essential for many speech-processing applications. Deep learning models have achieved high performance, but often fail when the training and infe…

eess.AS2024

UCIL: An Unsupervised Class Incremental Learning Approach for Sound Event Detection

Yang Xiao, Rohan Kumar Das

This work explores class-incremental learning (CIL) for sound event detection (SED), advancing adaptability towards real-world scenarios. CIL's success in domains like computer vis…