99 citations · 300 across the 18 of their papers we have counts for
5 papers · 1 filter
Instruction-ViT: Multi-Modal Prompts for Instruction Learning in ViT
Zhenxiang Xiao, Yuzhong Chen, Lu Zhang +14
Prompts have been proven to play a crucial role in large language models, and in recent years, vision models have also been using prompts to improve scalability for multiple downst…
LWSIS: LiDAR-guided Weakly Supervised Instance Segmentation for Autonomous Driving
Xiang Li, Junbo Yin, Botian Shi +3
Image instance segmentation is a fundamental research topic in autonomous driving, which is crucial for scene understanding and road safety. Advanced learning-based approaches ofte…
Towards Spatial Equilibrium Object Detection
Zhaohui Zheng, Yuming Chen, Qibin Hou +2
Semantic objects are unevenly distributed over images. In this paper, we study the spatial disequilibrium problem of modern object detectors and propose to quantify this ``spatial…
DTG-SSOD: Dense Teacher Guidance for Semi-Supervised Object Detection
Gang Li, Xiang Li, Yujie Wang +3
The Mean-Teacher (MT) scheme is widely adopted in semi-supervised object detection (SSOD). In MT, the sparse pseudo labels, offered by the final predictions of the teacher (e.g., a…
Knowledge Distillation for Object Detection via Rank Mimicking and Prediction-guided Feature Imitation
Gang Li, Xiang Li, Yujie Wang +3
Knowledge Distillation (KD) is a widely-used technology to inherit information from cumbersome teacher models to compact student models, consequently realizing model compression an…