activity
20242026
collaborators

5 papers

cs.CV2026

BiPrompt: Bilateral Prompt Optimization for Visual and Textual Debiasing in Vision-Language Models

Sunny Gupta, Shounak Das, Amit Sethi

Vision language foundation models such as CLIP exhibit impressive zero-shot generalization yet remain vulnerable to spurious correlations across visual and textual modalities. Exis…

cs.CV2025

IDAL: Improved Domain Adaptive Learning for Natural Images Dataset

Ravi Kant Gupta, Shounak Das, Amit Sethi

We present a novel approach for unsupervised domain adaptation (UDA) for natural images. A commonly-used objective for UDA schemes is to enhance domain alignment in representation…

cs.AI2025

FEDTAIL: Federated Long-Tailed Domain Generalization with Sharpness-Guided Gradient Matching

Sunny Gupta, Nikita Jangid, Shounak Das +1

Domain Generalization (DG) seeks to train models that perform reliably on unseen target domains without access to target data during training. While recent progress in smoothing th…

cs.CV2025

Scalable Whole Slide Image Representation Using K-Mean Clustering and Fisher Vector Aggregation

Ravi Kant Gupta, Shounak Das, Ardhendu Sekhar +1

Whole slide images (WSIs) are high-resolution, gigapixel sized images that pose significant computational challenges for traditional machine learning models due to their size and h…

eess.IV2024

Clustered Patch Embeddings for Permutation-Invariant Classification of Whole Slide Images

Ravi Kant Gupta, Shounak Das, Amit Sethi

Whole Slide Imaging (WSI) is a cornerstone of digital pathology, offering detailed insights critical for diagnosis and research. Yet, the gigapixel size of WSIs imposes significant…