1 paper
Chenyu Yang, Xizhou Zhu, Jinguo Zhu +9
Recently, vision model pre-training has evolved from relying on manually annotated datasets to leveraging large-scale, web-crawled image-text data. Despite these advances, there is…