49 citations · 70 across the 4 of their papers we have counts for
4 papers
Bootstrap Fine-Grained Vision-Language Alignment for Unified Zero-Shot Anomaly Localization
Hanqiu Deng, Zhaoxiang Zhang, Jinan Bao +1
Contrastive Language-Image Pre-training (CLIP) models have shown promising performance on zero-shot visual recognition tasks by learning visual representations under natural langua…
Fully Sparse Fusion for 3D Object Detection
Yingyan Li, Lue Fan, Yang Liu +4
Currently prevalent multimodal 3D detection methods are built upon LiDAR-based detectors that usually use dense Bird's-Eye-View (BEV) feature maps. However, the cost of such BEV fe…
Pointly-Supervised Panoptic Segmentation
Junsong Fan, Zhaoxiang Zhang, Tieniu Tan
In this paper, we propose a new approach to applying point-level annotations for weakly-supervised panoptic segmentation. Instead of the dense pixel-level labels used by fully supe…
Self-Supervised Predictive Learning: A Negative-Free Method for Sound Source Localization in Visual Scenes
Zengjie Song, Yuxi Wang, Junsong Fan +2
Sound source localization in visual scenes aims to localize objects emitting the sound in a given image. Recent works showing impressive localization performance typically rely on…