3 papers
cs.CV2024
The Solution for Language-Enhanced Image New Category Discovery
Haonan Xu, Dian Chao, Xiangyu Wu +2
Treating texts as images, combining prompts with textual labels for prompt tuning, and leveraging the alignment properties of CLIP have been successfully applied in zero-shot multi…
eess.IV2024
High-Resolution Image Translation Model Based on Grayscale Redefinition
Xixian Wu, Dian Chao, Yang Yang
Image-to-image translation is a technique that focuses on transferring images from one domain to another while maintaining the essential content representations. In recent years, i…
cs.CV2024
The Solution for the ICCV 2023 1st Scientific Figure Captioning Challenge
Dian Chao, Xin Song, Shupeng Zhong +4
In this paper, we propose a solution for improving the quality of captions generated for figures in papers. We adopt the approach of summarizing the textual content in the paper to…