1 paper
Jiahao Liu, Senhao Cao
Large-scale Vision-Language models have achieved remarkable results in various domains, such as image captioning and conditioned image generation. Nevertheless, these models still…