activity
20242026
collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

Spectral Tail Auxiliary Learning for AI-Generated Image Detection

Xingyi Li, Jiahui Zhang, Yiheng Li +2

As generative image models evolve rapidly, the perceptual gap between generated and real images continues to narrow, making AI-generated image detection increasingly challenging. M…

cs.CV2025

PiT: Progressive Diffusion Transformer

Jiafu Wu, Yabiao Wang, Jian Li +4

Diffusion Transformers (DiTs) achieve remarkable performance within image generation via the transformer architecture. Conventionally, DiTs are constructed by stacking serial isotr…

cs.CV2025

Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection

Yuhu Bai, Jiangning Zhang, Yunkang Cao +4

With the advent of vision-language models (e.g., CLIP) in zero- and few-shot settings, CLIP has been widely applied to zero-shot anomaly detection (ZSAD) in recent research, where…

cs.CV2024

Dual-path Frequency Discriminators for Few-shot Anomaly Detection

Yuhu Bai, Jiangning Zhang, Zhaofeng Chen +3

Few-shot anomaly detection (FSAD) plays a crucial role in industrial manufacturing. However, existing FSAD methods encounter difficulties leveraging a limited number of normal samp…

cs.CV2024

VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation

Qilin Wang, Zhengkai Jiang, Chengming Xu +7

Human image animation involves generating a video from a static image by following a specified pose sequence. Current approaches typically adopt a multi-stage pipeline that separat…