11 citations · 14 across the 3 of their papers we have counts for
3 papers
cs.CV2021★ 11 cited
Image Comes Dancing with Collaborative Parsing-Flow Video Synthesis
Bowen Wu, Zhenyu Xie, Xiaodan Liang +3
Transferring human motion from a source to a target person poses great potential in computer vision and graphics applications. A crucial step is to manipulate sequential future mot…
cs.CL2021
Wav-BERT: Cooperative Acoustic and Linguistic Representation Learning for Low-Resource Speech Recognition
Guolin Zheng, Yubei Xiao, Ke Gong +3
Unifying acoustic and linguistic representation learning has become increasingly crucial to transfer the knowledge learned on the abundance of high-resource language data for low-r…
cs.CL2020★ 3 cited
Adversarial Meta Sampling for Multilingual Low-Resource Speech Recognition
Yubei Xiao, Ke Gong, Pan Zhou +3
Low-resource automatic speech recognition (ASR) is challenging, as the low-resource target language data cannot well train an ASR model. To solve this issue, meta-learning formulat…