1 citations · 1 across the 5 of their papers we have counts for
6 papers
FlexSED: Towards Open-Vocabulary Sound Event Detection
Jiarui Hai, Helin Wang, Weizhe Guo +1
Despite recent progress in large-scale sound event detection (SED) systems capable of handling hundreds of sound classes, existing multi-class classification frameworks remain fund…
SynSonic: Augmenting Sound Event Detection through Text-to-Audio Diffusion ControlNet and Effective Sample Filtering
Jiarui Hai, Mounya Elhilali
Data synthesis and augmentation are essential for Sound Event Detection (SED) due to the scarcity of temporally labeled data. While augmentation methods like SpecAugment and Mix-up…
DreamVoice: Text-Guided Voice Conversion
Jiarui Hai, Karan Thakkar, Helin Wang +2
Generative voice technologies are rapidly evolving, offering opportunities for more personalized and inclusive experiences. Traditional one-shot voice conversion (VC) requires a ta…
Investigating Self-Supervised Deep Representations for EEG-based Auditory Attention Decoding
Karan Thakkar, Jiarui Hai, Mounya Elhilali
Auditory Attention Decoding (AAD) algorithms play a crucial role in isolating desired sound sources within challenging acoustic environments directly from brain activity. Although…
DPM-TSE: A Diffusion Probabilistic Model for Target Sound Extraction
Jiarui Hai, Helin Wang, Dongchao Yang +3
Common target sound extraction (TSE) approaches primarily relied on discriminative approaches in order to separate the target sound while minimizing interference from the unwanted…
Joint Acoustic and Class Inference for Weakly Supervised Sound Event Detection
Sandeep Kothinti, Keisuke Imoto, Debmalya Chakrabarty +3
Sound event detection is a challenging task, especially for scenes with multiple simultaneous events. While event classification methods tend to be fairly accurate, event localizat…