1 paper
Zhemin Zhang, Xun Gong
Foundation models, such as CNNs and ViTs, have powered the development of image representation learning. However, general guidance to model architecture design is still missing. In…