19 citations · 45 across the 13 of their papers we have counts for
20 papers · 1 filter
Video Object of Interest Segmentation
Siyuan Zhou, Chunru Zhan, Biao Wang +3
In this work, we present a new computer vision task named video object of interest segmentation (VOIS). Given a video and a target image of interest, our objective is to simultaneo…
Motion and Appearance Adaptation for Cross-Domain Motion Transfer
Borun Xu, Biao Wang, Jinhong Deng +5
Motion transfer aims to transfer the motion of a driving video to a source image. When there are considerable differences between object in the driving video and that in the source…
Motion Transformer for Unsupervised Image Animation
Jiale Tao, Biao Wang, Tiezheng Ge +3
Image animation aims to animate a source image by using motion learned from a driving video. Current state-of-the-art methods typically use convolutional neural networks (CNNs) to…
Dual-Level Decoupled Transformer for Video Captioning
Yiqi Gao, Xinglin Hou, Wei Suo +4
Video captioning aims to understand the spatio-temporal semantic concept of the video and generate descriptive sentences. The de-facto approach to this task dictates a text generat…
CapOnImage: Context-driven Dense-Captioning on Image
Yiqi Gao, Xinglin Hou, Yuanmeng Zhang +3
Existing image captioning systems are dedicated to generating narrative captions for images, which are spatially detached from the image in presentation. However, texts can also be…
Self-Supervised Text Erasing with Controllable Image Synthesis
Gangwei Jiang, Shiyao Wang, Tiezheng Ge +3
Recent efforts on scene text erasing have shown promising results. However, existing methods require rich yet costly label annotations to obtain robust models, which limits the use…