3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 1 cited
ParFormer: A Vision Transformer with Parallel Mixer and Sparse Channel Attention Patch Embedding
Novendra Setyawan, Ghufron Wahyu Kurniawan, Chi-Chia Sun +3
Convolutional Neural Networks (CNNs) and Transformers have achieved remarkable success in computer vision tasks. However, their deep architectures often lead to high computational…
cs.CV2024★ 3 cited
Forging Vision Foundation Models for Autonomous Driving: Challenges, Methodologies, and Opportunities
Xu Yan, Haiming Zhang, Yingjie Cai +13
The rise of large foundation models, trained on extensive datasets, is revolutionizing the field of AI. Models such as SAM, DALL-E2, and GPT-4 showcase their adaptability by extrac…