activity
20172023
most citedEVA-02: A Visual Representation for Neon Genesis

212 citations · 469 across the 13 of their papers we have counts for

collaborators
Showing 2022 · cs.CVShow all

5 papers · 2 filters

cs.CV2022★ 23 cited

EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

Yuxin Fang, Wen Wang, Binhui Xie +6

We launch EVA, a vision-centric foundation model to explore the limits of visual representation at scale using only publicly accessible data. EVA is a vanilla ViT pre-trained to re…

cs.CV2022★ 10 cited

Images Speak in Images: A Generalist Painter for In-Context Visual Learning

Xinlong Wang, Wen Wang, Yue Cao +2

In-context learning, as a new paradigm in NLP, allows the model to rapidly adapt to various tasks with only a handful of prompts and examples. But in computer vision, the difficult…

cs.CV2022★ 2 cited

Point-Teaching: Weakly Semi-Supervised Object Detection with Point Annotations

Yongtao Ge, Qiang Zhou, Xinlong Wang +3

Point annotations are considerably more time-efficient than bounding box annotations. However, how to use cheap point annotations to boost the performance of semi-supervised object…

cs.CV2022★ 3 cited

FreeSOLO: Learning to Segment Objects without Annotations

Xinlong Wang, Zhiding Yu, Shalini De Mello +4

Instance segmentation is a fundamental vision task that aims to recognize and segment each object in an image. However, it requires costly annotations such as bounding boxes and se…

cs.CV2022

Poseur: Direct Human Pose Regression with Transformers

Weian Mao, Yongtao Ge, Chunhua Shen +4

We propose a direct, regression-based approach to 2D human pose estimation from single images. We formulate the problem as a sequence prediction task, which we solve using a Transf…