14 citations · 84 across the 42 of their papers we have counts for
42 papers
AnyRefill: A Unified, Data-Efficient Framework for Left-Prompt-Guided Vision Tasks
Ming Xie, Chenjie Cao, Yunuo Cai +3
In this paper, we present a novel Left-Prompt-Guided (LPG) paradigm to address a diverse range of reference-based vision tasks. Inspired by the human creative process, we reformula…
DuMo: Dual Encoder Modulation Network for Precise Concept Erasure
Feng Han, Kai Chen, Chao Gong +3
The exceptional generative capability of text-to-image models has raised substantial safety concerns regarding the generation of Not-Safe-For-Work (NSFW) content and potential copy…
AIM: Additional Image Guided Generation of Transferable Adversarial Attacks
Teng Li, Xingjun Ma, Yu-Gang Jiang
Transferable adversarial examples highlight the vulnerability of deep neural networks (DNNs) to imperceptible perturbations across various real-world applications. While there have…
STNMamba: Mamba-based Spatial-Temporal Normality Learning for Video Anomaly Detection
Zhangxun Li, Mengyang Zhao, Xuan Yang +6
Video anomaly detection (VAD) has been extensively researched due to its potential for intelligent video systems. However, most existing methods based on CNNs and transformers stil…
Comprehensive Multi-Modal Prototypes are Simple and Effective Classifiers for Vast-Vocabulary Object Detection
Yitong Chen, Wenhao Yao, Lingchen Meng +3
Enabling models to recognize vast open-world categories has been a longstanding pursuit in object detection. By leveraging the generalization capabilities of vision-language models…
Retrieval Augmented Recipe Generation
Guoshan Liu, Hailong Yin, Bin Zhu +3
Given the potential applications of generating recipes from food images, this area has garnered significant attention from researchers in recent years. Existing works for recipe ge…