◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xiaodi Wang

12 papers hereh-index 10862 citations22 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author12

Across the 12 of 12 papers where every author was matched, so the position is known.

fields
  • cs.CV12
same name
  • Xiaodi Wang — 1 paper, h 6
  • Xiaodi Wang — 1 paper, h 2
  • Xiaodi Wang — 1 paper, h 1
  • Xiaodi Wang — 1 paper, h 1
  • Xiaodi Wang — 1 paper, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20182025
most citedOriented Object Detection with Transformer

32 citations · 80 across the 10 of their papers we have counts for

collaborators
Showing 2022 · cs.CVShow all

4 papers · 2 filters

cs.CV2022★ 8 cited

CAE v2: Context Autoencoder with CLIP Target

Xinyu Zhang, Jiahui Chen, Junkun Yuan +10

Masked image modeling (MIM) learns visual representation by masking and reconstructing image patches. Applying the reconstruction supervision on the CLIP representation has been pr…

cs.CV2022★ 19 cited

Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining

Qiang Chen, Jian Wang, Chuchu Han +12

We present a strong object detector with encoder-decoder pretraining and finetuning. Our method, called Group DETR v2, is built upon a vision transformer encoder ViT-Huge~\cite{dos…

cs.CV2022★ 4 cited

MAFormer: A Transformer Network with Multi-scale Attention Fusion for Visual Recognition

Yunhao Wang, Huixin Sun, Xiaodi Wang +6

Vision Transformer and its variants have demonstrated great potential in various computer vision tasks. But conventional vision transformers often focus on global dependency at a c…

cs.CV2022★ 9 cited

Context Autoencoder for Self-Supervised Representation Learning

Xiaokang Chen, Mingyu Ding, Xiaodi Wang +7

We present a novel masked image modeling (MIM) approach, context autoencoder (CAE), for self-supervised representation pretraining. We pretrain an encoder by making predictions in…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.