52 citations · 121 across the 5 of their papers we have counts for
5 papers · 1 filter
IDa-Det: An Information Discrepancy-aware Distillation for 1-bit Detectors
Sheng Xu, Yanjing Li, Bohan Zeng +5
Knowledge distillation (KD) has been proven to be useful for training compact object detection models. However, we observe that KD is often effective when the teacher model and stu…
ConvMAE: Masked Convolution Meets Masked Autoencoders
Peng Gao, Teli Ma, Hongsheng Li +3
Vision Transformers (ViT) become widely-adopted architectures for various vision tasks. Masked auto-encoding for feature pretraining and multi-scale hybrid convolution-transformer…
TerViT: An Efficient Ternary Vision Transformer
Sheng Xu, Yanjing Li, Teli Ma +4
Vision transformers (ViTs) have demonstrated great potential in various visual tasks, but suffer from expensive computational and memory cost problems when deployed on resource-con…
Oriented Object Detection with Transformer
Teli Ma, Mingyuan Mao, Honghui Zheng +6
Object detection with Transformers (DETR) has achieved a competitive performance over traditional detectors, such as Faster R-CNN. However, the potential of DETR remains largely un…
Dual-stream Network for Visual Recognition
Mingyuan Mao, Renrui Zhang, Honghui Zheng +6
Transformers with remarkable global representation capacities achieve competitive results for visual tasks, but fail to consider high-level local pattern information in input image…