activity
20172022
most citedVision-to-Language Tasks Based on Attributes and Attention Mechanism

53 citations · 196 across the 10 of their papers we have counts for

collaborators

11 papers

cs.CV202216 cited

Pairwise Comparison Network for Remote Sensing Scene Classification

Zhang Yue, Zheng Xiangtao, Lu Xiaoqiang

Remote sensing scene classification aims to assign a specific semantic label to a remote sensing image. Recently, convolutional neural networks have greatly improved the performanc…

cs.CV2021

Reconstructive Sequence-Graph Network for Video Summarization

Bin Zhao, Haopeng Li, Xiaoqiang Lu +1

Exploiting the inner-shot and inter-shot dependencies is essential for key-shot based video summarization. Current approaches mainly devote to modeling the video as a frame sequenc…

cs.SD202124 cited

Audio Description from Image by Modal Translation Network

Hailong Ning, Xiangtao Zheng, Yuan Yuan +1

Audio is the main form for the visually impaired to obtain information. In reality, all kinds of visual data always exist, but audio data does not exist in many cases. In order to…

cs.CV202129 cited

Bio-Inspired Representation Learning for Visual Attention Prediction

Yuan Yuan, Hailong Ning, Xiaoqiang Lu

Visual Attention Prediction (VAP) is a significant and imperative issue in the field of computer vision. Most of existing VAP methods are based on deep learning. However, they do n…

cs.CV201953 cited

Vision-to-Language Tasks Based on Attributes and Attention Mechanism

Xuelong Li, Aihong Yuan, Xiaoqiang Lu

Vision-to-language tasks aim to integrate computer vision and natural language processing together, which has attracted the attention of many researchers. For typical approaches, t…

cs.CV20194 cited

Hierarchical Recurrent Neural Network for Video Summarization

Bin Zhao, Xuelong Li, Xiaoqiang Lu

Exploiting the temporal dependency among video frames or subshots is very important for the task of video summarization. Practically, RNN is good at temporal dependency modeling, a…