5 papers · 1 filter
Spectral Tail Auxiliary Learning for AI-Generated Image Detection
Xingyi Li, Jiahui Zhang, Yiheng Li +2
As generative image models evolve rapidly, the perceptual gap between generated and real images continues to narrow, making AI-generated image detection increasingly challenging. M…
PiT: Progressive Diffusion Transformer
Jiafu Wu, Yabiao Wang, Jian Li +4
Diffusion Transformers (DiTs) achieve remarkable performance within image generation via the transformer architecture. Conventionally, DiTs are constructed by stacking serial isotr…
Bridge Feature Matching and Cross-Modal Alignment with Mutual-filtering for Zero-shot Anomaly Detection
Yuhu Bai, Jiangning Zhang, Yunkang Cao +4
With the advent of vision-language models (e.g., CLIP) in zero- and few-shot settings, CLIP has been widely applied to zero-shot anomaly detection (ZSAD) in recent research, where…
Dual-path Frequency Discriminators for Few-shot Anomaly Detection
Yuhu Bai, Jiangning Zhang, Zhaofeng Chen +3
Few-shot anomaly detection (FSAD) plays a crucial role in industrial manufacturing. However, existing FSAD methods encounter difficulties leveraging a limited number of normal samp…
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
Qilin Wang, Zhengkai Jiang, Chengming Xu +7
Human image animation involves generating a video from a static image by following a specified pose sequence. Current approaches typically adopt a multi-stage pipeline that separat…