1 citations · 1 across the 2 of their papers we have counts for
4 papers
A Semantic Information-based Hierarchical Speech Enhancement Method Using Factorized Codec and Diffusion Model
Yang Xiang, Canan Huang, Desheng Hu +3
Most current speech enhancement (SE) methods recover clean speech from noisy inputs by directly estimating time-frequency masks or spectrums. However, these approaches often neglec…
The RoyalFlush Automatic Speech Diarization and Recognition System for In-Car Multi-Channel Automatic Speech Recognition Challenge
Jingguang Tian, Shuaishuai Ye, Shunfei Chen +4
This paper presents our system submission for the In-Car Multi-Channel Automatic Speech Recognition (ICMC-ASR) Challenge, which focuses on speaker diarization and speech recognitio…
A Deep Representation Learning-based Speech Enhancement Method Using Complex Convolution Recurrent Variational Autoencoder
Yang Xiang, Jingguang Tian, Xinhui Hu +2
Generally, the performance of deep neural networks (DNNs) heavily depends on the quality of data representation learning. Our preliminary work has emphasized the significance of de…
Large-Scale Learning on Overlapped Speech Detection: New Benchmark and New General System
Zhaohui Yin, Jingguang Tian, Xinhui Hu +2
Overlapped Speech Detection (OSD) is an important part of speech applications involving analysis of multi-party conversations. However, most of existing OSD systems are trained and…