2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CV2025
Omni-AD: Learning to Reconstruct Global and Local Features for Multi-class Anomaly Detection
Jiajie Quan, Ao Tong, Yuxuan Cai +3
In multi-class unsupervised anomaly detection(MUAD), reconstruction-based methods learn to map input images to normal patterns to identify anomalous pixels. However, this strategy…
cs.CV2024
LLaVA-KD: A Framework of Distilling Multimodal Large Language Models
Yuxuan Cai, Jiangning Zhang, Haoyang He +7
The success of Large Language Models (LLMs) has inspired the development of Multimodal Large Language Models (MLLMs) for unified understanding of vision and language. However, the…
cs.CV2024★ 2 cited
Anomaly Detection by Adapting a pre-trained Vision Language Model
Yuxuan Cai, Xinwei He, Dingkang Liang +2
Recently, large vision and language models have shown their success when adapting them to many downstream tasks. In this paper, we present a unified framework named CLIP-ADA for An…