10 citations · 10 across the 3 of their papers we have counts for
4 papers
CLIP4Caption: CLIP for Video Caption
Mingkang Tang, Zhanyu Wang, Zhenhua Liu +3
Video captioning is a challenging task since it requires generating sentences describing various diverse and complex videos. Existing video captioning models lack adequate visual r…
A Coarse-to-Fine Instance Segmentation Network with Learning Boundary Representation
Feng Luo, Bin-Bin Gao, Jiangpeng Yan +1
Boundary-based instance segmentation has drawn much attention since of its attractive efficiency. However, existing methods suffer from the difficulty in long-distance regression.…
Disentangled Non-Local Neural Networks
Minghao Yin, Zhuliang Yao, Yue Cao +4
The non-local block is a popular module for strengthening the context modeling ability of a regular convolutional neural network. This paper first studies the non-local block in de…
Structure from Recurrent Motion: From Rigidity to Recurrency
Xiu Li, Hongdong Li, Hanbyul Joo +2
This paper proposes a new method for Non-Rigid Structure-from-Motion (NRSfM) from a long monocular video sequence observing a non-rigid object performing recurrent and possibly rep…