1 paper
Zhenyuan Lu
To improve performance in visual feature representation from photos or videos for practical applications, we generally require large-scale human-annotated labeled data while traini…