activity
20242026
most citedHyperCap: Hyperspectral Land Cover Captioning Dataset for Vision Language Models

1 citations · 1 across the 4 of their papers we have counts for

collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV20261 cited

HyperCap: Hyperspectral Land Cover Captioning Dataset for Vision Language Models

Aryan Das, Tanishq Rachamalla, Pravendra Singh +5

We introduce HyperCap, the first large-scale hyperspectral captioning dataset designed to enhance model performance and effectiveness in remote sensing applications. Unlike traditi…

cs.CV2026

Efficient Text-Guided Convolutional Adapter for the Diffusion Model

Aryan Das, Koushik Biswas, Swalpa Kumar Roy +2

We introduce the Nexus Adapters, novel text-guided efficient adapters to the diffusion-based framework for the Structure Preserving Conditional Generation (SPCG). Recently, structu…

cs.CV2026

Uncertainty-Aware Vision-Language Segmentation for Medical Imaging

Aryan Das, Tanishq Rachamalla, Koushik Biswas +2

We introduce a novel uncertainty-aware multimodal segmentation framework that leverages both radiological images and associated clinical text for precise medical diagnosis. We prop…

cs.CV2025

SceneMixer: Exploring Convolutional Mixing Networks for Remote Sensing Scene Classification

Mohammed Q. Alkhatib, Ali Jamali, Swalpa Kumar Roy

Remote sensing scene classification plays a key role in Earth observation by enabling the automatic identification of land use and land cover (LULC) patterns from aerial and satell…

cs.CV2025

TD-RD: A Top-Down Benchmark with Real-Time Framework for Road Damage Detection

Xi Xiao, Zhengji Li, Wentao Wang +5

Object detection has witnessed remarkable advancements over the past decade, largely driven by breakthroughs in deep learning and the proliferation of large scale datasets. However…

cs.CV2024

Spatial and Spatial-Spectral Morphological Mamba for Hyperspectral Image Classification

Muhammad Ahmad, Muhammad Hassaan Farooq Butt, Adil Mehmood Khan +6

Recent advancements in transformers, specifically self-attention mechanisms, have significantly improved hyperspectral image (HSI) classification. However, these models often suffe…