2 papers
cs.CV2026
MMLGNet: Cross-Modal Alignment of Remote Sensing Data using CLIP
Aditya Chaudhary, Sneha Barman, Mainak Singha +3
In this paper, we propose a novel multimodal framework, Multimodal Language-Guided Network (MMLGNet), to align heterogeneous remote sensing modalities like Hyperspectral Imaging (H…
cs.CV2025
Two-Stage Vision Transformer for Image Restoration: Colorization Pretraining + Residual Upsampling
Aditya Chaudhary, Prachet Dev Singh, Ankit Jha
In computer vision, Single Image Super-Resolution (SISR) is still a difficult problem. We present ViT-SR, a new technique to improve the performance of a Vision Transformer (ViT) e…