activity
20222024
most citedDAB-DETR: Dynamic Anchor Boxes are Better Queries for DETR

400 citations · 562 across the 23 of their papers we have counts for

collaborators
Showing cs.CVShow all

12 papers · 1 filter

cs.CV20242 cited

Spatio-Temporal Side Tuning Pre-trained Foundation Models for Video-based Pedestrian Attribute Recognition

Xiao Wang, Qian Zhu, Jiandong Jin +5

Existing pedestrian attribute recognition (PAR) algorithms are mainly developed based on a static image, however, the performance is unreliable in challenging scenarios, such as he…

cs.CV20239 cited

DPM-Solver-v3: Improved Diffusion ODE Solver with Empirical Model Statistics

Kaiwen Zheng, Cheng Lu, Jianfei Chen +1

Diffusion probabilistic models (DPMs) have exhibited excellent performance for high-fidelity image generation while suffering from inefficient sampling. Recent works accelerate the…

cs.CV20234 cited

How Robust is Google's Bard to Adversarial Image Attacks?

Yinpeng Dong, Huanran Chen, Jiawei Chen +6

Multimodal Large Language Models (MLLMs) that integrate text and other modalities (especially vision) have achieved unprecedented performance in various multimodal tasks. However,…

cs.CV2023

PREIM3D: 3D Consistent Precise Image Attribute Editing from a Single Image

Jianhui Li, Jianmin Li, Haoji Zhang +5

We study the 3D-aware image attribute editing problem in this paper, which has wide applications in practice. Recent methods solved the problem by training a shared encoder to map…

cs.CV20233 cited

Learning CLIP Guided Visual-Text Fusion Transformer for Video-based Pedestrian Attribute Recognition

Jun Zhu, Jiandong Jin, Zihan Yang +2

Existing pedestrian attribute recognition (PAR) algorithms are mainly developed based on a static image. However, the performance is not reliable for images with challenging factor…

cs.CV20233 cited

A Closer Look at Parameter-Efficient Tuning in Diffusion Models

Chendong Xiang, Fan Bao, Chongxuan Li +2

Large-scale diffusion models like Stable Diffusion are powerful and find various real-world applications while customizing such models by fine-tuning is both memory and time ineffi…