7 papers
Naming the Concepts Classifiers Rely On: Language-Anchored Decomposition for Faithful Explanation
Ahsan Habib Akash, Dipkamal Bhusal, Stacey Jones +3
Deep neural networks are widely deployed in high-stakes visual applications where interpretability is critical, yet existing explanations face a trade-off: post-hoc concept methods…
Two is better than one: A Collapse-free Multi-Reward RLIF Training Framework
Shourov Joarder, Diganta Sikdar, Ahsan Habib Akash +2
Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of LLMs, but often depends on external supervision from human annotations or…
Addressing Bias in VLMs for Glaucoma Detection Without Protected Attribute Supervision
Ahsan Habib Akash, Greg Murray, Annahita Amireskandari +4
Vision-Language Models (VLMs) have achieved remarkable success on multimodal tasks such as image-text retrieval and zero-shot classification, yet they can exhibit demographic biase…
SW-ViT: A Spatio-Temporal Vision Transformer Network with Post Denoiser for Sequential Multi-Push Ultrasound Shear Wave Elastography
Ahsan Habib Akash, MD Jahin Alam, Md. Kamrul Hasan
Objective: Ultrasound Shear Wave Elastography (SWE) demonstrates great potential in assessing soft-tissue pathology by mapping tissue stiffness, which is linked to malignancy. Trad…
Robust CNN Multi-Nested-LSTM Framework with Compound Loss for Patch-based Multi-Push Ultrasound Shear Wave Imaging and Segmentation
Md. Jahin Alam, Ahsan Habib, Md. Kamrul Hasan
Ultrasound Shear Wave Elastography (SWE) is a noteworthy tool for in-vivo noninvasive tissue pathology assessment. State-of-the-art techniques can generate reasonable estimates of…
SegCodeNet: Color-Coded Segmentation Masks for Activity Detection from Wearable Cameras
Asif Shahriyar Sushmit, Partho Ghosh, Md. Abrar Istiak +3
Activity detection from first-person videos (FPV) captured using a wearable camera is an active research field with potential applications in many sectors, including healthcare, la…