61 citations · 123 across the 12 of their papers we have counts for
21 papers · 1 filter
Training-free Transformer Architecture Search
Qinqin Zhou, Kekai Sheng, Xiawu Zheng +5
Recently, Vision Transformer (ViT) has achieved remarkable success in several computer vision tasks. The progresses are highly relevant to the architecture design, then it is worth…
RMNet: Equivalently Removing Residual Connection from Networks
Fanxu Meng, Hao Cheng, Jiaxin Zhuang +2
Although residual connection enables training very deep neural networks, it is not friendly for online inference due to its multi-branch topology. This encourages many researchers…
PR-Net: Preference Reasoning for Personalized Video Highlight Detection
Runnan Chen, Penghao Zhou, Wenzhe Wang +4
Personalized video highlight detection aims to shorten a long video to interesting moments according to a user's preference, which has recently raised the community's attention. Cu…
Learning Canonical View Representation for 3D Shape Recognition with Arbitrary Views
Xin Wei, Yifei Gong, Fudong Wang +2
In this paper, we focus on recognizing 3D shapes from arbitrary views, i.e., arbitrary numbers and positions of viewpoints. It is a challenging and realistic setting for view-based…
Updatable Siamese Tracker with Two-stage One-shot Learning
Xinglong Sun, Guangliang Han, Lihong Guo +3
Offline Siamese networks have achieved very promising tracking performance, especially in accuracy and efficiency. However, they often fail to track an object in complex scenes due…
Temporal Modulation Network for Controllable Space-Time Video Super-Resolution
Gang Xu, Jun Xu, Zhen Li +3
Space-time video super-resolution (STVSR) aims to increase the spatial and temporal resolutions of low-resolution and low-frame-rate videos. Recently, deformable convolution based…