652 citations · 1k across the 2 of their papers we have counts for
Showing 2014 · cs.CVShow all
2 papers · 2 filters
cs.CV2014★ 652 cited
Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)
Junhua Mao, Wei Xu, Yi Yang +3
In this paper, we present a multimodal Recurrent Neural Network (m-RNN) model for generating novel image captions. It directly models the probability distribution of generating a w…
cs.CV2014★ 370 cited
Explain Images with Multimodal Recurrent Neural Networks
Junhua Mao, Wei Xu, Yi Yang +2
In this paper, we present a multimodal Recurrent Neural Network (m-RNN) model for generating novel sentence descriptions to explain the content of images. It directly models the pr…