18 citations · 25 across the 6 of their papers we have counts for
4 papers · 1 filter
TerraFM: A Scalable Foundation Model for Unified Multisensor Earth Observation
Muhammad Sohail Danish, Muhammad Akhtar Munir, Syed Roshaan Ali Shah +5
Modern Earth observation (EO) increasingly leverages deep learning to harness the scale and diversity of satellite imagery across sensors and regions. While recent foundation model…
O-TPT: Orthogonality Constraints for Calibrating Test-time Prompt Tuning in Vision-Language Models
Ashshak Sharifdeen, Muhammad Akhtar Munir, Sanoojan Baliah +2
Test-time prompt tuning for vision-language models (VLMs) is getting attention because of their ability to learn with unlabeled data without fine-tuning. Although test-time prompt…
Cal-DETR: Calibrated Detection Transformer
Muhammad Akhtar Munir, Salman Khan, Muhammad Haris Khan +2
Albeit revealing impressive predictive performance for several computer vision tasks, deep neural networks (DNNs) are prone to making overconfident predictions. This limits the ado…
Video Instance Segmentation via Multi-scale Spatio-temporal Split Attention Transformer
Omkar Thawakar, Sanath Narayan, Jiale Cao +6
State-of-the-art transformer-based video instance segmentation (VIS) approaches typically utilize either single-scale spatio-temporal features or per-frame multi-scale features dur…