3 papers
cs.CV2024
TG-LMM: Enhancing Medical Image Segmentation Accuracy through Text-Guided Large Multi-Modal Model
Yihao Zhao, Enhao Zhong, Cuiyun Yuan +5
We propose TG-LMM (Text-Guided Large Multi-Modal Model), a novel approach that leverages textual descriptions of organs to enhance segmentation accuracy in medical images. Existing…
cs.CV2020
Unpaired Image-to-Image Translation using Adversarial Consistency Loss
Yihao Zhao, Ruihai Wu, Hao Dong
Unpaired image-to-image translation is a class of vision problems whose goal is to find the mapping between different image domains using unpaired training data. Cycle-consistency…
cs.CV2019
DLGAN: Disentangling Label-Specific Fine-Grained Features for Image Manipulation
Guanqi Zhan, Yihao Zhao, Bingchan Zhao +3
Recent studies have shown how disentangling images into content and feature spaces can provide controllable image translation/ manipulation. In this paper, we propose a framework t…