2 papers
cs.CV2025
Demystify Transformers & Convolutions in Modern Image Deep Networks
Xiaowei Hu, Min Shi, Weiyun Wang +9
Vision transformers have gained popularity recently, leading to the development of new vision backbones with improved features and consistent performance gains. However, these adva…
cs.CV2024
NVDS+: Towards Efficient and Versatile Neural Stabilizer for Video Depth Estimation
Yiran Wang, Min Shi, Jiaqi Li +7
Video depth estimation aims to infer temporally consistent depth. One approach is to finetune a single-image model on each video with geometry constraints, which proves inefficient…