3 papers
cs.CV2024
UniTT-Stereo: Unified Training of Transformer for Enhanced Stereo Matching
Soomin Kim, Hyesong Choi, Jihye Ahn +1
Unlike other vision tasks where Transformer-based approaches are becoming increasingly common, stereo depth estimation is still dominated by convolution-based approaches. This is m…
cs.CV2024
SG-MIM: Structured Knowledge Guided Efficient Pre-training for Dense Prediction
Sumin Son, Hyesong Choi, Dongbo Min
Masked Image Modeling (MIM) techniques have redefined the landscape of computer vision, enabling pre-trained models to achieve exceptional performance across a broad spectrum of ta…
cs.CV2024
Dynamic Guidance Adversarial Distillation with Enhanced Teacher Knowledge
Hyejin Park, Dongbo Min
In the realm of Adversarial Distillation (AD), strategic and precise knowledge transfer from an adversarially robust teacher model to a less robust student model is paramount. Our…