348 citations · 799 across the 15 of their papers we have counts for
44 papers · 1 filter
Temporal Modulation Network for Controllable Space-Time Video Super-Resolution
Gang Xu, Jun Xu, Zhen Li +3
Space-time video super-resolution (STVSR) aims to increase the spatial and temporal resolutions of low-resolution and low-frame-rate videos. Recently, deformable convolution based…
Decoupled Spatial Temporal Graphs for Generic Visual Grounding
Qianyu Feng, Yunchao Wei, Mingming Cheng +1
Visual grounding is a long-lasting problem in vision-language understanding due to its diversity and complexity. Current practices concentrate mostly on performing visual grounding…
Global2Local: Efficient Structure Search for Video Action Segmentation
Shang-Hua Gao, Qi Han, Zhong-Yu Li +3
Temporal receptive fields of models play an important role in action segmentation. Large receptive fields facilitate the long-term relations among video clips while small receptive…
Centralized Information Interaction for Salient Object Detection
Jiang-Jiang Liu, Zhi-Ang Liu, Ming-Ming Cheng
The U-shape structure has shown its advantage in salient object detection for efficiently combining multi-scale features. However, most existing U-shape based methods focused on im…
Dense Attention Fluid Network for Salient Object Detection in Optical Remote Sensing Images
Qijian Zhang, Runmin Cong, Chongyi Li +5
Despite the remarkable advances in visual saliency analysis for natural scene images (NSIs), salient object detection (SOD) for optical remote sensing images (RSIs) still remains a…
Delving Deep into Label Smoothing
Chang-Bin Zhang, Peng-Tao Jiang, Qibin Hou +4
Label smoothing is an effective regularization tool for deep neural networks (DNNs), which generates soft labels by applying a weighted average between the uniform distribution and…