36 citations · 42 across the 9 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2025
Audiobook-CC: Controllable Long-context Speech Generation for Multicast Audiobook
Min Liu, JingJing Yin, Xiang Zhang +6
Existing text-to-speech systems predominantly focus on single-sentence synthesis and lack adequate contextual modeling as well as fine-grained performance control capabilities for…
eess.AS2021★ 1 cited
Poformer: A simple pooling transformer for speaker verification
Yufeng Ma, Yiwei Ding, Miao Zhao +3
Most recent speaker verification systems are based on extracting speaker embeddings using a deep neural network. The pooling layer in the network aims to aggregate frame-level feat…
eess.AS2018
End-to-end Speech Recognition with Adaptive Computation Steps
Mohan Li, Min Liu, Masanori Hattori
In this paper, we present Adaptive Computation Steps (ACS) algo-rithm, which enables end-to-end speech recognition models to dy-namically decide how many frames should be processed…