118 citations · 194 across the 6 of their papers we have counts for
4 papers · 1 filter
Expanding Language-Image Pretrained Models for General Video Recognition
Bolin Ni, Houwen Peng, Minghao Chen +5
Contrastive language-image pretraining has shown great success in learning visual-textual joint representation from web-scale data, demonstrating remarkable "zero-shot" generalizat…
Multi-Content Complementation Network for Salient Object Detection in Optical Remote Sensing Images
Gongyang Li, Zhi Liu, Weisi Lin +1
In the computer vision community, great progresses have been achieved in salient object detection from natural scene images (NSI-SOD); by contrast, salient object detection in opti…
PFLD: A Practical Facial Landmark Detector
Xiaojie Guo, Siyuan Li, Jinke Yu +5
Being accurate, efficient, and compact is essential to a facial landmark detector for practical use. To simultaneously consider the three concerns, this paper investigates a neat m…
Multi-level Contextual RNNs with Attention Model for Scene Labeling
Heng Fan, Xue Mei, Danil Prokhorov +1
Context in image is crucial for scene labeling while existing methods only exploit local context generated from a small surrounding area of an image patch or a pixel, by contrast l…