7 citations · 9 across the 3 of their papers we have counts for
4 papers
MR-SVS: Singing Voice Synthesis with Multi-Reference Encoder
Shoutong Wang, Jinglin Liu, Yi Ren +3
Multi-speaker singing voice synthesis is to generate the singing voice sung by different speakers. To generalize to new speakers, previous zero-shot singing adaptation methods obta…
WSRGlow: A Glow-based Waveform Generative Model for Audio Super-Resolution
Kexun Zhang, Yi Ren, Changliang Xu +1
Audio super-resolution is the task of constructing a high-resolution (HR) audio from a low-resolution (LR) audio by adding the missing band. Previous methods based on convolutional…
SoccerDB: A Large-Scale Database for Comprehensive Video Understanding
Yudong Jiang, Kaixu Cui, Leilei Chen +2
Soccer videos can serve as a perfect research object for video understanding because soccer games are played under well-defined rules while complex and intriguing enough for resear…
Comprehensive Video Understanding: Video summarization with content-based video recommender design
Yudong Jiang, Kaixu Cui, Bo Peng +1
Video summarization aims to extract keyframes/shots from a long video. Previous methods mainly take diversity and representativeness of generated summaries as prior knowledge in al…