5 citations · 13 across the 5 of their papers we have counts for
5 papers
TF-Blender: Temporal Feature Blender for Video Object Detection
Yiming Cui, Liqi Yan, Zhiwen Cao +1
Video objection detection is a challenging task because isolated video frames may encounter appearance deterioration, which introduces great confusion for detection. One of the pop…
Hierarchical Attention Fusion for Geo-Localization
Liqi Yan, Yiming Cui, Yingjie Chen +1
Geo-localization is a critical task in computer vision. In this work, we cast the geo-localization as a 2D image retrieval task. Current state-of-the-art methods for 2D geo-localiz…
DenserNet: Weakly Supervised Visual Localization Using Multi-scale Feature Aggregation
Dongfang Liu, Yiming Cui, Liqi Yan +3
In this work, we introduce a Denser Feature Network (DenserNet) for visual localization. Our work provides three principal contributions. First, we develop a convolutional neural n…
Multimodal Aggregation Approach for Memory Vision-Voice Indoor Navigation with Meta-Learning
Liqi Yan, Dongfang Liu, Yaoxian Song +1
Vision and voice are two vital keys for agents' interaction and learning. In this paper, we present a novel indoor navigation model called Memory Vision-Voice Indoor Navigation (MV…
Crowd Video Captioning
Liqi Yan, Mingjian Zhu, Changbin Yu
Describing a video automatically with natural language is a challenging task in the area of computer vision. In most cases, the on-site situation of great events is reported in new…