3 papers
cs.CV2024
Multi-Modality Driven LoRA for Adverse Condition Depth Estimation
Guanglei Yang, Rui Tian, Yongqiang Zhang +3
The autonomous driving community is increasingly focused on addressing corner case problems, particularly those related to ensuring driving safety under adverse conditions (e.g., n…
cs.CV2022
ResFormer: Scaling ViTs with Multi-Resolution Training
Rui Tian, Zuxuan Wu, Qi Dai +3
Vision Transformers (ViTs) have achieved overwhelming success, yet they suffer from vulnerable resolution scalability, i.e., the performance drops drastically when presented with i…
cs.CV2022
Deeper Insights into the Robustness of ViTs towards Common Corruptions
Rui Tian, Zuxuan Wu, Qi Dai +2
With Vision Transformers (ViTs) making great advances in a variety of computer vision tasks, recent literature have proposed various variants of vanilla ViTs to achieve better effi…