1 paper · 1 filter
Hao Tang, Chenwei Xie, Xiaoyi Bao +4
In this paper, we propose UniLIP, a unified framework that adapts CLIP for multimodal understanding, generation and editing. Although CLIP excels at understanding, it lacks reconst…