27 citations · 41 across the 17 of their papers we have counts for
6 papers · 1 filter
RCT: Random Consistency Training for Semi-supervised Sound Event Detection
Nian Shao, Erfan Loweimi, Xiaofei Li
Sound event detection (SED), as a core module of acoustic environmental analysis, suffers from the problem of data deficiency. The integration of semi-supervised learning (SSL) lar…
Multi-channel Narrow-band Deep Speech Separation with Full-band Permutation Invariant Training
Changsheng Quan, Xiaofei Li
This paper addresses the problem of multi-channel multi-speech separation based on deep learning techniques. In the short time Fourier transform domain, we propose an end-to-end na…
AcousticFusion: Fusing Sound Source Localization to Visual SLAM in Dynamic Environments
Tianwei Zhang, Huayan Zhang, Xiaofei Li +3
Dynamic objects in the environment, such as people and other agents, lead to challenges for existing simultaneous localization and mapping (SLAM) approaches. To deal with dynamic e…
Microphone Array Generalization for Multichannel Narrowband Deep Speech Enhancement
Siyuan Zhang, Xiaofei Li
This paper addresses the problem of microphone array generalization for deep-learning-based end-to-end multichannel speech enhancement. We aim to train a unique deep neural network…
SizeNet: Object Recognition via Object Real Size-based Convolutional Networks
Xiaofei Li, Zhong Dong
Inspired by the conclusion that humans choose the visual cortex regions corresponding to the real size of an object to analyze its features when identifying objects in the real wor…
Semi-supervised Sound Event Detection using Random Augmentation and Consistency Regularization
Xiaofei Li
Sound event detection is a core module for acoustic environmental analysis. Semi-supervised learning technique allows to largely scale up the dataset without increasing the annotat…