41 citations · 82 across the 15 of their papers we have counts for
17 papers · 1 filter
Ground4D: Consistency-Aware 4D Reconstruction from Monocular Video
Qing Zhao, Weijian Deng, Pengxu Wei +1
Learning a 4D scene representation from a single monocular video that supports dynamic novel-view synthesis while maintaining faithful geometry over time remains challenging. Dynam…
Where Detectors Fail: Probing Generative Space for Generalizable AI-Generated Image Detection
Zijie Cao, Weijie Tu, Yao Xiao +3
Detecting AI-generated images (AIGI) remains challenging because detectors often fail to generalize to unseen generators. Although existing methods are trained on large datasets, t…
Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention into Convolutions
ZiYi Dong, Chengxing Zhou, Weijian Deng +3
Contemporary diffusion models built upon U-Net or Diffusion Transformer (DiT) architectures have revolutionized image generation through transformer-based attention mechanisms. The…
Decoder-Only LLMs are Better Controllers for Diffusion Models
Ziyi Dong, Yao Xiao, Pengxu Wei +1
Groundbreaking advancements in text-to-image generation have recently been achieved with the emergence of diffusion models. These models exhibit a remarkable ability to generate hi…
SAM-COD: SAM-guided Unified Framework for Weakly-Supervised Camouflaged Object Detection
Huafeng Chen, Pengxu Wei, Guangqian Guo +1
Most Camouflaged Object Detection (COD) methods heavily rely on mask annotations, which are time-consuming and labor-intensive to acquire. Existing weakly-supervised COD approaches…
Diff-Mosaic: Augmenting Realistic Representations in Infrared Small Target Detection via Diffusion Prior
Yukai Shi, Yupei Lin, Pengxu Wei +3
Recently, researchers have proposed various deep learning methods to accurately detect infrared targets with the characteristics of indistinct shape and texture. Due to the limited…