7 papers
Referring Camouflaged Object Detection
Xuying Zhang, Bowen Yin, Zheng Lin +3
We consider the problem of referring camouflaged object detection (Ref-COD), a new task that aims to segment specified camouflaged objects based on a small set of referring images…
YOLO-MS: Rethinking Multi-Scale Representation Learning for Real-time Object Detection
Yuming Chen, Xinbin Yuan, Jiabao Wang +4
We aim at providing the object detection community with an efficient and performant object detector, termed YOLO-MS. The core design is based on a series of investigations on how m…
StyleDiffusion: Prompt-Embedding Inversion for Text-Based Editing
Senmao Li, Joost van de Weijer, Taihang Hu +5
A significant research effort is focused on exploiting the amazing capacities of pretrained diffusion models for the editing of images.They either finetune the model, or invert the…
SRFormerV2: Taking a Closer Look at Permuted Self-Attention for Image Super-Resolution
Yupeng Zhou, Zhen Li, Chun-Le Guo +3
Previous works have shown that increasing the window size for Transformer-based image super-resolution models (e.g., SwinIR) can significantly improve the model performance. Still,…
Zone Evaluation: Revealing Spatial Bias in Object Detection
Zhaohui Zheng, Yuming Chen, Qibin Hou +3
A fundamental limitation of object detectors is that they suffer from "spatial bias", and in particular perform less satisfactorily when detecting objects near image borders. For a…
StoryDiffusion: Consistent Self-Attention for Long-Range Image and Video Generation
Yupeng Zhou, Daquan Zhou, Ming-Ming Cheng +2
For recent diffusion-based generative models, maintaining consistent content across a series of generated images, especially those containing subjects and complex details, presents…