4 citations · 9 across the 3 of their papers we have counts for
4 papers
CNN-based Discriminative Training for Domain Compensation in Acoustic Event Detection with Frame-wise Classifier
Tiantian Tang, Xinyuan Zhou, Yanhua Long +2
Domain mismatch is a noteworthy issue in acoustic event detection tasks, as the target domain data is difficult to access in most real applications. In this study, we propose a nov…
Multi-channel target speech extraction with channel decorrelation and target speaker adaptation
Jiangyu Han, Xinyuan Zhou, Yanhua Long +1
The end-to-end approaches for single-channel target speech extraction have attracted widespread attention. However, the studies for end-to-end multi-channel target speech extractio…
Multi-Encoder-Decoder Transformer for Code-Switching Speech Recognition
Xinyuan Zhou, Emre Yılmaz, Yanhua Long +2
Code-switching (CS) occurs when a speaker alternates words of two or more languages within a single sentence or across sentences. Automatic speech recognition (ASR) of CS speech ha…
Self-and-Mixed Attention Decoder with Deep Acoustic Structure for Transformer-based LVCSR
Xinyuan Zhou, Grandee Lee, Emre Yılmaz +3
The Transformer has shown impressive performance in automatic speech recognition. It uses the encoder-decoder structure with self-attention to learn the relationship between the hi…