652 citations · 1.1k across the 3 of their papers we have counts for
3 papers
cs.CV2014★ 652 cited
Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)
Junhua Mao, Wei Xu, Yi Yang +3
In this paper, we present a multimodal Recurrent Neural Network (m-RNN) model for generating novel image captions. It directly models the probability distribution of generating a w…
cs.CV2014★ 370 cited
Explain Images with Multimodal Recurrent Neural Networks
Junhua Mao, Wei Xu, Yi Yang +2
In this paper, we present a multimodal Recurrent Neural Network (m-RNN) model for generating novel sentence descriptions to explain the content of images. It directly models the pr…
cs.CV2014★ 95 cited
Learning Fine-grained Image Similarity with Deep Ranking
Jiang Wang, Yang song, Thomas Leung +5
Learning fine-grained image similarity is a challenging task. It needs to capture between-class and within-class image differences. This paper proposes a deep ranking model that em…