400 citations · 562 across the 23 of their papers we have counts for
12 papers · 1 filter
Spatio-Temporal Side Tuning Pre-trained Foundation Models for Video-based Pedestrian Attribute Recognition
Xiao Wang, Qian Zhu, Jiandong Jin +5
Existing pedestrian attribute recognition (PAR) algorithms are mainly developed based on a static image, however, the performance is unreliable in challenging scenarios, such as he…
DPM-Solver-v3: Improved Diffusion ODE Solver with Empirical Model Statistics
Kaiwen Zheng, Cheng Lu, Jianfei Chen +1
Diffusion probabilistic models (DPMs) have exhibited excellent performance for high-fidelity image generation while suffering from inefficient sampling. Recent works accelerate the…
How Robust is Google's Bard to Adversarial Image Attacks?
Yinpeng Dong, Huanran Chen, Jiawei Chen +6
Multimodal Large Language Models (MLLMs) that integrate text and other modalities (especially vision) have achieved unprecedented performance in various multimodal tasks. However,…
PREIM3D: 3D Consistent Precise Image Attribute Editing from a Single Image
Jianhui Li, Jianmin Li, Haoji Zhang +5
We study the 3D-aware image attribute editing problem in this paper, which has wide applications in practice. Recent methods solved the problem by training a shared encoder to map…
Learning CLIP Guided Visual-Text Fusion Transformer for Video-based Pedestrian Attribute Recognition
Jun Zhu, Jiandong Jin, Zihan Yang +2
Existing pedestrian attribute recognition (PAR) algorithms are mainly developed based on a static image. However, the performance is not reliable for images with challenging factor…
A Closer Look at Parameter-Efficient Tuning in Diffusion Models
Chendong Xiang, Fan Bao, Chongxuan Li +2
Large-scale diffusion models like Stable Diffusion are powerful and find various real-world applications while customizing such models by fine-tuning is both memory and time ineffi…