13 papers
MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection
Xiaobin Hu, Haoyang He, Bo Yin +5
While recent advancements in anomaly detection have demonstrated the efficacy of CNN- and Transformer-based approaches, these architectures face inherent limitations: CNNs struggle…
Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval
Zhucun Xue, Jiangning Zhang, Juntao Jiang +6
The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult to synthesize. Traditional biblio…
OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing
Haoyang He, Jie Wang, Jiangning Zhang +5
The quality and diversity of instruction-based image editing datasets are continuously increasing, yet large-scale, high-quality datasets for instruction-based video editing remain…
JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing
Xiaohu Huang, Hao Zhou, Haoyang He +3
In this paper, we present JoVA, a streamlined framework that unifies joint video-audio generation and editing. While existing methods often rely on fragmented, task-specific archit…
EfficientIML: Efficient High-Resolution Image Manipulation Localization
Jinhan Li, Haoyang He, Lei Xie +1
With imaging devices delivering ever-higher resolutions and the emerging diffusion-based forgery methods, current detectors trained only on traditional datasets (with splicing, cop…
A Comprehensive Library for Benchmarking Multi-class Visual Anomaly Detection
Jiangning Zhang, Haoyang He, Zhenye Gan +7
Visual anomaly detection aims to identify anomalous regions in images through unsupervised learning paradigms, with increasing application demand and value in fields such as indust…