57 citations · 78 across the 17 of their papers we have counts for
13 papers · 1 filter
Dynamic Video Frame Interpolation with integrated Difficulty Pre-Assessment
Ban Chen, Xin Jin, Youxin Chen +4
Video frame interpolation(VFI) has witnessed great progress in recent years. While existing VFI models still struggle to achieve a good trade-off between accuracy and efficiency: f…
[CLS] Token is All You Need for Zero-Shot Semantic Segmentation
Letian Wu, Wenyao Zhang, Tengping Jiang +3
In this paper, we propose an embarrassingly simple yet highly effective zero-shot semantic segmentation (ZS3) method, based on the pre-trained vision-language model CLIP. First, ou…
TBFormer: Two-Branch Transformer for Image Forgery Localization
Yaqi Liu, Binbin Lv, Xin Jin +2
Image forgery localization aims to identify forged regions by capturing subtle traces from high-quality discriminative features. In this paper, we propose a Transformer-style netwo…
Aesthetics Driven Autonomous Time-Lapse Photography Generation by Virtual and Real Robots
Xiaobo Gao, Qi Kuang, Xin Jin +3
Time-lapse photography is employed in movies and promotional films because it can reflect the passage of time in a short time and strengthen the visual attraction. However, since i…
Aesthetic Visual Question Answering of Photographs
Xin Jin, Wu Zhou, Xinghui Zhou +4
Aesthetic assessment of images can be categorized into two main forms: numerical assessment and language assessment. Aesthetics caption of photographs is the only task of aesthetic…
Aesthetic Language Guidance Generation of Images Using Attribute Comparison
Xin Jin, Qiang Deng, Jianwen Lv +3
With the vigorous development of mobile photography technology, major mobile phone manufacturers are scrambling to improve the shooting ability of equipments and the photo beautifi…