288 citations · 294 across the 4 of their papers we have counts for
4 papers · 1 filter
Towards Small Object Editing: A Benchmark Dataset and A Training-Free Approach
Qihe Pan, Zhen Zhao, Zicheng Wang +5
A plethora of text-guided image editing methods has recently been developed by leveraging the impressive capabilities of large-scale diffusion-based generative models especially St…
SOEDiff: Efficient Distillation for Small Object Editing
Yiming Wu, Qihe Pan, Zhen Zhao +3
In this paper, we delve into a new task known as small object editing (SOE), which focuses on text-based image inpainting within a constrained, small-sized area. Despite the remark…
Training-Free Unsupervised Prompt for Vision-Language Models
Sifan Long, Linbin Wang, Zhen Zhao +4
Prompt learning has become the most effective paradigm for adapting large pre-trained vision-language models (VLMs) to downstream tasks. Recently, unsupervised prompt tuning method…
HAP: Structure-Aware Masked Image Modeling for Human-Centric Perception
Junkun Yuan, Xinyu Zhang, Hao Zhou +12
Model pre-training is essential in human-centric perception. In this paper, we first introduce masked image modeling (MIM) as a pre-training approach for this task. Upon revisiting…