activity
20162022
most citedContextual Non-Local Alignment over Full-Scale Representation for Text-Based Person Search

61 citations · 123 across the 12 of their papers we have counts for

collaborators
Showing cs.CVShow all

21 papers · 1 filter

cs.CV2022

Training-free Transformer Architecture Search

Qinqin Zhou, Kekai Sheng, Xiawu Zheng +5

Recently, Vision Transformer (ViT) has achieved remarkable success in several computer vision tasks. The progresses are highly relevant to the architecture design, then it is worth…

cs.CV20219 cited

RMNet: Equivalently Removing Residual Connection from Networks

Fanxu Meng, Hao Cheng, Jiaxin Zhuang +2

Although residual connection enables training very deep neural networks, it is not friendly for online inference due to its multi-branch topology. This encourages many researchers…

cs.CV2021

PR-Net: Preference Reasoning for Personalized Video Highlight Detection

Runnan Chen, Penghao Zhou, Wenzhe Wang +4

Personalized video highlight detection aims to shorten a long video to interesting moments according to a user's preference, which has recently raised the community's attention. Cu…

cs.CV20211 cited

Learning Canonical View Representation for 3D Shape Recognition with Arbitrary Views

Xin Wei, Yifei Gong, Fudong Wang +2

In this paper, we focus on recognizing 3D shapes from arbitrary views, i.e., arbitrary numbers and positions of viewpoints. It is a challenging and realistic setting for view-based…

cs.CV2021

Updatable Siamese Tracker with Two-stage One-shot Learning

Xinglong Sun, Guangliang Han, Lihong Guo +3

Offline Siamese networks have achieved very promising tracking performance, especially in accuracy and efficiency. However, they often fail to track an object in complex scenes due…

cs.CV2021

Temporal Modulation Network for Controllable Space-Time Video Super-Resolution

Gang Xu, Jun Xu, Zhen Li +3

Space-time video super-resolution (STVSR) aims to increase the spatial and temporal resolutions of low-resolution and low-frame-rate videos. Recently, deformable convolution based…