activity
20242026
collaborators

13 papers

cs.CV2026

MambaADv2: Evolving Duality-enhanced State Space Model for Unsupervised Anomaly Detection

Xiaobin Hu, Haoyang He, Bo Yin +5

While recent advancements in anomaly detection have demonstrated the efficacy of CNN- and Transformer-based approaches, these architectures face inherent limitations: CNNs struggle…

cs.CV2026

Multi-Dimensional Knowledge Profiling with Large-Scale Literature Database and Hierarchical Retrieval

Zhucun Xue, Jiangning Zhang, Juntao Jiang +6

The rapid expansion of research across machine learning, vision, and language has produced a volume of publications that is increasingly difficult to synthesize. Traditional biblio…

cs.CV2025

OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Haoyang He, Jie Wang, Jiangning Zhang +5

The quality and diversity of instruction-based image editing datasets are continuously increasing, yet large-scale, high-quality datasets for instruction-based video editing remain…

cs.CV2025

JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing

Xiaohu Huang, Hao Zhou, Haoyang He +3

In this paper, we present JoVA, a streamlined framework that unifies joint video-audio generation and editing. While existing methods often rely on fragmented, task-specific archit…

cs.CV2025

EfficientIML: Efficient High-Resolution Image Manipulation Localization

Jinhan Li, Haoyang He, Lei Xie +1

With imaging devices delivering ever-higher resolutions and the emerging diffusion-based forgery methods, current detectors trained only on traditional datasets (with splicing, cop…

cs.CV2025

A Comprehensive Library for Benchmarking Multi-class Visual Anomaly Detection

Jiangning Zhang, Haoyang He, Zhenye Gan +7

Visual anomaly detection aims to identify anomalous regions in images through unsupervised learning paradigms, with increasing application demand and value in fields such as indust…