18 papers
Benchmarking the Domain Gap: Model Selection Instability Under Domain Shift in Video Capsule Endoscopy
Dan Hanson, Debesh Jha
Video capsule endoscopy (VCE) classification is typically evaluated within a single dataset, yet clinical deployment demands robustness across acquisition sources, labeling policie…
Induce to Empower: Improving Lightweight Baselines via Foundation Model Induction for Generalized Polyp Segmentation
Shivanshu Agnihotri, Snehashis Majhi, Deepak Ranjan Nayak +2
Automated polyp segmentation in colonoscopy continues to pose challenges due to substantial appearance variations and indistinct polyp boundaries. Although emerging foundation mode…
DentiAsk: A VQA Benchmark for Multimodal Reasoning in Panoramic Dental Radiographs
Debesh Jha, Tapas Kumar Dutta, Roshan Paudel +9
Accurate interpretation of panoramic dental radiographs requires the integration of multiple reasoning capabilities: detection, spatial localization, and quantitative assessment. D…
SRMA-Mamba: Spatial Reverse Mamba Attention Network for Pathological Liver Segmentation in MRI Volumes
Jun Zeng, Quoc-Huy Trinh, Deepak Ranjan Nayak +3
Liver cirrhosis plays a critical role in the prognosis of chronic liver disease. Early detection and timely intervention are essential for reducing mortality rates. However, the in…
Revisiting LLM Adaptation for 3D CT Report Generation: A Study of Scaling and Diagnostic Priors
Vanshali Sharma, Andrea M. Bejar, Halil Ertugrul Aktas +4
Recent advances in multimodal learning, including large language models (LLMs) and vision-language models (VLMs), have demonstrated strong adaptability to natural images. However,…
PRS-Med: Position Reasoning Segmentation in Medical Imaging
Quoc-Huy Trinh, Minh-Van Nguyen, Jun Zeng +2
Prompt-based medical image segmentation has rapidly emerged, yet existing methods rely on explicit prompts like bounding boxes and struggle to reason about the spatial relationship…