3 papers
cs.CV2026
A-TPT: Angular Diversity Calibration Properties for Test-Time Prompt Tuning of Vision-Language Models
Shihab Aaqil Ahamed, Udaya S. K. P. Miriya Thanthrige, Ranga Rodrigo +1
Test-time prompt tuning (TPT) has emerged as a promising technique for adapting large vision-language models (VLMs) to unseen tasks without relying on labeled data. However, the la…
cs.CV2025
SENCA-st: Integrating Spatial Transcriptomics and Histopathology with Cross Attention Shared Encoder for Region Identification in Cancer Pathology
Shanaka Liyanaarachchi, Chathurya Wijethunga, Shihab Aaqil Ahamed +2
Spatial transcriptomics is an emerging field that enables the identification of functional regions based on the spatial distribution of gene expression. Integrating this functional…
cs.CV2025
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
Shihab Aaqil Ahamed, Malitha Gunawardhana, Liel David +3
Current video-based Masked Autoencoders (MAEs) primarily focus on learning effective spatiotemporal representations from a visual perspective, which may lead the model to prioritiz…