158 citations · 426 across the 20 of their papers we have counts for
Showing 2017Show all
3 papers · 1 filter
cs.CV2017★ 16 cited
From Deterministic to Generative: Multi-Modal Stochastic RNNs for Video Captioning
Jingkuan Song, Yuyu Guo, Lianli Gao +3
Video captioning in essential is a complex natural process, which is affected by various uncertainties stemming from video content, subjective judgment, etc. In this paper we build…
cs.CV2017★ 36 cited
Hierarchical LSTM with Adjusted Temporal Attention for Video Captioning
Jingkuan Song, Zhao Guo, Lianli Gao +3
Recent progress has been made in using attention based encoder-decoder framework for video captioning. However, most existing decoders apply the attention mechanism to every genera…
cs.CV2017★ 32 cited
Deep Region Hashing for Efficient Large-scale Instance Search from Images
Jingkuan Song, Tao He, Lianli Gao +2
Instance Search (INS) is a fundamental problem for many applications, while it is more challenging comparing to traditional image search since the relevancy is defined at the insta…