1 citations · 1 across the 6 of their papers we have counts for
6 papers
GeoStack: A Framework for Quasi-Abelian Knowledge Composition in VLMs
Pranav Mantini, Shishir K. Shah
We address the challenge of knowledge composition in Vision-Language Models (VLMs), where accumulating expertise across multiple domains or tasks typically leads to catastrophic fo…
BiCLIP: Domain Canonicalization via Structured Geometric Transformation
Pranav Mantini, Shishir K. Shah
Recent advances in vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities, yet adapting these models to specialized domains remains a significant challen…
CCPA: Long-term Person Re-Identification via Contrastive Clothing and Pose Augmentation
Vuong D. Nguyen, Shishir K. Shah
Long-term Person Re-Identification (LRe-ID) aims at matching an individual across cameras after a long period of time, presenting variations in clothing, pose, and viewpoint. In th…
Data Quality Aware Approaches for Addressing Model Drift of Semantic Segmentation Models
Samiha Mirza, Vuong D. Nguyen, Pranav Mantini +1
In the midst of the rapid integration of artificial intelligence (AI) into real world applications, one pressing challenge we confront is the phenomenon of model drift, wherein the…
Attention-based Shape and Gait Representations Learning for Video-based Cloth-Changing Person Re-Identification
Vuong D. Nguyen, Samiha Mirza, Pranav Mantini +1
Current state-of-the-art Video-based Person Re-Identification (Re-ID) primarily relies on appearance features extracted by deep learning models. These methods are not applicable fo…
A Survey of Feature Types and Their Contributions for Camera Tampering Detection
Pranav Mantini, Shishir K. Shah
Camera tamper detection is the ability to detect unauthorized and unintentional alterations in surveillance cameras by analyzing the video. Camera tampering can occur due to natura…