3 papers
eess.AS2025
SEF-PNet: Speaker Encoder-Free Personalized Speech Enhancement with Local and Global Contexts Aggregation
Ziling Huang, Haixin Guan, Haoran Wei +1
Personalized speech enhancement (PSE) methods typically rely on pre-trained speaker verification models or self-designed speaker encoders to extract target speaker clues, guiding t…
cs.SD2023
Autoencoder with Group-based Decoder and Multi-task Optimization for Anomalous Sound Detection
Yifan Zhou, Dongxing Xu, Haoran Wei +1
In industry, machine anomalous sound detection (ASD) is in great demand. However, collecting enough abnormal samples is difficult due to the high cost, which boosts the rapid devel…
cs.SD2023
Multi-pass Training and Cross-information Fusion for Low-resource End-to-end Accented Speech Recognition
Xuefei Wang, Yanhua Long, Yijie Li +1
Low-resource accented speech recognition is one of the important challenges faced by current ASR technology in practical applications. In this study, we propose a Conformer-based a…