234 citations · 396 across the 8 of their papers we have counts for
14 papers · 1 filter
Rethinking the Number of Shots in Robust Model-Agnostic Meta-Learning
Xiaoyue Duan, Guoliang Kang, Runqi Wang +4
Robust Model-Agnostic Meta-Learning (MAML) is usually adopted to train a meta-model which may fast adapt to novel classes with only a few exemplars and meanwhile remain robust to a…
CAE v2: Context Autoencoder with CLIP Target
Xinyu Zhang, Jiahui Chen, Junkun Yuan +10
Masked image modeling (MIM) learns visual representation by masking and reconstructing image patches. Applying the reconstruction supervision on the CLIP representation has been pr…
Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining
Qiang Chen, Jian Wang, Chuchu Han +12
We present a strong object detector with encoder-decoder pretraining and finetuning. Our method, called Group DETR v2, is built upon a vision transformer encoder ViT-Huge~\cite{dos…
Oriented Object Detection with Transformer
Teli Ma, Mingyuan Mao, Honghui Zheng +6
Object detection with Transformers (DETR) has achieved a competitive performance over traditional detectors, such as Faster R-CNN. However, the potential of DETR remains largely un…
Probabilistic Ranking-Aware Ensembles for Enhanced Object Detections
Mingyuan Mao, Baochang Zhang, David Doermann +5
Model ensembles are becoming one of the most effective approaches for improving object detection performance already optimized for a single detector. Conventional methods directly…
Dual-stream Network for Visual Recognition
Mingyuan Mao, Renrui Zhang, Honghui Zheng +6
Transformers with remarkable global representation capacities achieve competitive results for visual tasks, but fail to consider high-level local pattern information in input image…