25 citations · 57 across the 12 of their papers we have counts for
12 papers
ClipSAM: CLIP and SAM Collaboration for Zero-Shot Anomaly Segmentation
Shengze Li, Jianjian Cao, Peng Ye +3
Recently, foundational models such as CLIP and SAM have shown promising performance for the task of Zero-Shot Anomaly Segmentation (ZSAS). However, either CLIP-based or SAM-based Z…
SpVOS: Efficient Video Object Segmentation with Triple Sparse Convolution
Weihao Lin, Tao Chen, Chong Yu
Semi-supervised video object segmentation (Semi-VOS), which requires only annotating the first frame of a video to segment future frames, has received increased attention recently.…
VQ-NeRF: Vector Quantization Enhances Implicit Neural Representations
Yiying Yang, Wen Liu, Fukun Yin +4
Recent advancements in implicit neural representations have contributed to high-fidelity surface reconstruction and photorealistic novel view synthesis. However, the computational…
Rethinking Cross-Domain Pedestrian Detection: A Background-Focused Distribution Alignment Framework for Instance-Free One-Stage Detectors
Yancheng Cai, Bo Zhang, Baopu Li +4
Cross-domain pedestrian detection aims to generalize pedestrian detectors from one label-rich domain to another label-scarce domain, which is crucial for various real-world applica…
Experts Weights Averaging: A New General Training Scheme for Vision Transformers
Yongqi Huang, Peng Ye, Xiaoshui Huang +4
Structural re-parameterization is a general training scheme for Convolutional Neural Networks (CNNs), which achieves performance improvement without increasing inference cost. As V…
Attention Consistency Refined Masked Frequency Forgery Representation for Generalizing Face Forgery Detection
Decheng Liu, Tao Chen, Chunlei Peng +3
Due to the successful development of deep image generation technology, visual data forgery detection would play a more important role in social and economic security. Existing forg…