7 citations · 10 across the 4 of their papers we have counts for
5 papers
Contextual Expressive Text-to-Speech
Jianhong Tu, Zeyu Cui, Xiaohuan Zhou +4
The goal of expressive Text-to-speech (TTS) is to synthesize natural speech with desired content, prosody, emotion, or timbre, in high expressiveness. Most of previous studies atte…
ViBERTgrid: A Jointly Trained Multi-Modal 2D Document Representation for Key Information Extraction from Documents
Weihong Lin, Qifang Gao, Lei Sun +4
Recent grid-based document representations like BERTgrid allow the simultaneous encoding of the textual and layout information of a document in a 2D feature map so that state-of-th…
Meta Feature Modulator for Long-tailed Recognition
Renzhen Wang, Kaiqin Hu, Yanwen Zhu +3
Deep neural networks often degrade significantly when training data suffer from class imbalance problems. Existing approaches, e.g., re-sampling and re-weighting, commonly address…
RotationOut as a Regularization Method for Neural Network
Kai Hu, Barnabas Poczos
In this paper, we propose a novel regularization method, RotationOut, for neural networks. Different from Dropout that handles each neuron/channel independently, RotationOut regard…
Qiniu Submission to ActivityNet Challenge 2018
Xiaoteng Zhang, Yixin Bao, Feiyun Zhang +7
In this paper, we introduce our submissions for the tasks of trimmed activity recognition (Kinetics) and trimmed event recognition (Moments in Time) for Activitynet Challenge 2018.…