27 citations · 35 across the 4 of their papers we have counts for
4 papers · 1 filter
Enhance Image-to-Image Generation with LLaVA-generated Prompts
Zhicheng Ding, Panfeng Li, Qikai Yang +1
This paper presents a novel approach to enhance image-to-image generation by leveraging the multimodal capabilities of the Large Language and Vision Assistant (LLaVA). We propose a…
Regional Style and Color Transfer
Zhicheng Ding, Panfeng Li, Qikai Yang +2
This paper presents a novel contribution to the field of regional style transfer. Existing methods often suffer from the drawback of applying style homogeneously across the entire…
CLIP as RNN: Segment Countless Visual Concepts without Training Endeavor
Shuyang Sun, Runjia Li, Philip Torr +2
Existing open-vocabulary image segmentation methods require a fine-tuning step on mask labels and/or image-text datasets. Mask labels are labor-intensive, which limits the number o…
Confidence Trigger Detection: Accelerating Real-time Tracking-by-detection Systems
Zhicheng Ding, Zhixin Lai, Siyang Li +3
Real-time object tracking necessitates a delicate balance between speed and accuracy, a challenge exacerbated by the computational demands of deep learning methods. In this paper,…