Showing eess.ASShow all
3 papers · 1 filter
eess.AS2023
Consistent and Relevant: Rethink the Query Embedding in General Sound Separation
Yuanyuan Wang, Hangting Chen, Dongchao Yang +4
The query-based audio separation usually employs specific queries to extract target sources from a mixture of audio signals. Currently, most query-based separation models need addi…
eess.AS2023
AutoPrep: An Automatic Preprocessing Framework for In-the-Wild Speech Data
Jianwei Yu, Hangting Chen, Yanyao Bian +6
Recently, the utilization of extensive open-sourced text data has significantly advanced the performance of text-based large language models (LLMs). However, the use of in-the-wild…
eess.AS2023
Complexity Scaling for Speech Denoising
Hangting Chen, Jianwei Yu, Chao Weng
Computational complexity is critical when deploying deep learning-based speech denoising models for on-device applications. Most prior research focused on optimizing model architec…