3 papers
cs.CL2024
ZALM3: Zero-Shot Enhancement of Vision-Language Alignment via In-Context Information in Multi-Turn Multimodal Medical Dialogue
Zhangpu Li, Changhong Zou, Suxue Ma +13
The rocketing prosperity of large language models (LLMs) in recent years has boosted the prevalence of vision-language models (VLMs) in the medical sector. In our online medical co…
cs.CV2024
SSL: A Self-similarity Loss for Improving Generative Image Super-resolution
Du Chen, Zhengqiang Zhang, Jie Liang +1
Generative adversarial networks (GAN) and generative diffusion models (DM) have been widely used in real-world image super-resolution (Real-ISR) to enhance the image perceptual qua…
cs.CV2023
Human Guided Ground-truth Generation for Realistic Image Super-resolution
Du Chen, Jie Liang, Xindong Zhang +3
How to generate the ground-truth (GT) image is a critical issue for training realistic image super-resolution (Real-ISR) models. Existing methods mostly take a set of high-resoluti…