1 paper
Zhiyu Zhao, Bingkun Huang, Sen Xing +3
Self-supervised foundation models have shown great potential in computer vision thanks to the pre-training paradigm of masked autoencoding. Scale is a primary factor influencing th…