36 citations · 39 across the 4 of their papers we have counts for
6 papers
The MSXF TTS System for ICASSP 2022 ADD Challenge
Chunyong Yang, Pengfei Liu, Yanli Chen +2
This paper presents our MSXF TTS system for Task 3.1 of the Audio Deep Synthesis Detection (ADD) Challenge 2022. We use an end to end text to speech system, and add a constraint lo…
Rep Works in Speaker Verification
Yufeng Ma, Miao Zhao, Yiwei Ding +3
Multi-branch convolutional neural network architecture has raised lots of attention in speaker verification since the aggregation of multiple parallel branches can significantly im…
Multi-query multi-head attention pooling and Inter-topK penalty for speaker verification
Miao Zhao, Yufeng Ma, Yiwei Ding +3
This paper describes the multi-query multi-head attention (MQMHA) pooling and inter-topK penalty methods which were first proposed in our submitted system description for VoxCeleb…
Poformer: A simple pooling transformer for speaker verification
Yufeng Ma, Yiwei Ding, Miao Zhao +3
Most recent speaker verification systems are based on extracting speaker embeddings using a deep neural network. The pooling layer in the network aims to aggregate frame-level feat…
The SpeakIn System for VoxCeleb Speaker Recognition Challange 2021
Miao Zhao, Yufeng Ma, Min Liu +1
This report describes our submission to the track 1 and track 2 of the VoxCeleb Speaker Recognition Challenge 2021 (VoxSRC 2021). Both track 1 and track 2 share the same speaker ve…
End-to-end Speech Recognition with Adaptive Computation Steps
Mohan Li, Min Liu, Masanori Hattori
In this paper, we present Adaptive Computation Steps (ACS) algo-rithm, which enables end-to-end speech recognition models to dy-namically decide how many frames should be processed…