84 citations · 275 across the 49 of their papers we have counts for
Showing 2018Show all
3 papers · 1 filter
cs.CV2018★ 17 cited
Hierarchical LSTMs with Adaptive Attention for Visual Captioning
Jingkuan Song, Xiangpeng Li, Lianli Gao +1
Recent progress has been made in using attention based encoder-decoder framework for image and video captioning. Most existing decoders apply the attention mechanism to every gener…
cs.CV2018
Neighbourhood Watch: Referring Expression Comprehension via Language-guided Graph Attention Networks
Peng Wang, Qi Wu, Jiewei Cao +3
The task in referring expression comprehension is to localise the object instance in an image described by a referring expression phrased in natural language. As a language-to-visi…
cs.CV2018
Self-Supervised Video Hashing with Hierarchical Binary Auto-encoder
Jingkuan Song, Hanwang Zhang, Xiangpeng Li +3
Existing video hash functions are built on three isolated stages: frame pooling, relaxed learning, and binarization, which have not adequately explored the temporal order of video…