42 citations · 137 across the 22 of their papers we have counts for
4 papers
Contextual Transformer Networks for Visual Recognition
Yehao Li, Ting Yao, Yingwei Pan +1
Transformer with self-attention has led to the revolutionizing of natural language processing field, and recently inspires the emergence of Transformer-style architecture design wi…
Deep Quantization: Encoding Convolutional Activations with Deep Generative Model
Zhaofan Qiu, Ting Yao, Tao Mei
Deep convolutional neural networks (CNNs) have proven highly effective for visual recognition, where learning a universal representation from activations of convolutional layer pla…
Video Captioning with Transferred Semantic Attributes
Yingwei Pan, Ting Yao, Houqiang Li +1
Automatically generating natural language descriptions of videos plays a fundamental challenge for computer vision community. Most recent progress in this problem has been achieved…
Boosting Image Captioning with Attributes
Ting Yao, Yingwei Pan, Yehao Li +2
Automatically describing an image with a natural language has been an emerging challenge in both fields of computer vision and natural language processing. In this paper, we presen…