Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
A Study of Finetuning Video Transformers for Multi-view Geometry Tasks
Huimin Wu, Kwang-Ting Cheng, Stephen Lin +1
This paper presents an investigation of vision transformer learning for multi-view geometry tasks, such as optical flow estimation, by fine-tuning video foundation models. Unlike p…
cs.CV2023
Exploring Transferability for Randomized Smoothing
Kai Qiu, Huishuai Zhang, Zhirong Wu +1
Training foundation models on extensive datasets and then finetuning them on specific tasks has emerged as the mainstream approach in artificial intelligence. However, the model ro…