activity
20242026
most citedHyperCap: Hyperspectral Land Cover Captioning Dataset for Vision Language Models

2 citations · 2 across the 6 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

OptiModNet: A UNet-Transformer Hybrid with Grouped-Query and Channel Attention for Optic Disc and Cup Segmentation

Soumili Ghosh, Debapriya Roy, Aryan Das +1

Precise segmentation of the optic disc and cup is critical for the early detection and diagnosis of glaucoma. However, achieving consistently high performance across datasets while…

cs.CV2026

UNITY: Attention Flow Networks for Adaptive Conditioning in Diffusion

Aryan Das, Koushik Biswas, Moloud Abdar +1

We introduce UNITY, a Universal-to-Specialized adapter for efficient and scalable composite conditioning in diffusion based image generation. Unlike prior methods that train separa…

cs.CV2026

Efficient Text-Guided Convolutional Adapter for the Diffusion Model

Aryan Das, Koushik Biswas, Swalpa Kumar Roy +2

We introduce the Nexus Adapters, novel text-guided efficient adapters to the diffusion-based framework for the Structure Preserving Conditional Generation (SPCG). Recently, structu…

cs.CV2026

Uncertainty-Aware Vision-Language Segmentation for Medical Imaging

Aryan Das, Tanishq Rachamalla, Koushik Biswas +2

We introduce a novel uncertainty-aware multimodal segmentation framework that leverages both radiological images and associated clinical text for precise medical diagnosis. We prop…

cs.CV20252 cited

HyperCap: Hyperspectral Land Cover Captioning Dataset for Vision Language Models

Aryan Das, Tanishq Rachamalla, Pravendra Singh +5

We introduce HyperCap, the first large-scale hyperspectral captioning dataset designed to enhance model performance and effectiveness in remote sensing applications. Unlike traditi…