25 papers
Recursive Vision Language Models for General Symbolic Reasoning
Omid Nejati Manzari, Guillaume Lajoie, Hassan Rivaz
Hard symbolic-reasoning tasks such as Sudoku, maze pathfinding, and ARC remain challenging for LLMs due to their fixed-depth autoregressive reasoning, which limits systematic searc…
Longitudinal Lesion Inpainting in Brain MRI via 3D Region Aware Diffusion
Zahra Karimaghaloo, Dumitru Fetco, Haz-Edine Assemlal +2
Accurate longitudinal analysis of brain MRI is often hindered by evolving lesions, which bias automated neuroimaging pipelines. While deep generative models have shown promise in i…
Evi-Steer: Learning to Steer Biomedical Vision-Language Models through Efficient and Generalizable Evidential Tuning
Taha Koleilat, Hassan Rivaz, Yiming Xiao
Parameter-efficient adaptation of vision-language foundation models is crucial for precise multimodal understanding of biomedical images, yet existing methods remain deterministic…
VesselSim: learning 3D blood vessel segmentation without expert annotations
Erin Rainville, Melissa Ananian, Tristan Mirolla +2
Blood vessel segmentation is a core task in medical image analysis for the care of vascular diseases and surgical planning, yet the challenges of providing expert vascular annotati…
Lightweight Physics-Aware Zero-Shot Ultrasound Plane-Wave Denoising
Hojat Asgariandehkordi, Mostafa Sharifzadeh, Morteza Rezanejad +1
Ultrasound Coherent Plane-Wave Compounding (CPWC) enhances image contrast by combining echoes from multiple steered transmissions. While increasing the number of steering angles ge…
CLIP-SVD: Efficient and Interpretable Vision-Language Adaptation via Singular Values
Taha Koleilat, Hassan Rivaz, Yiming Xiao
Vision-language models (VLMs) like CLIP have shown impressive zero-shot and few-shot learning capabilities across diverse applications. However, adapting these models to new fine-g…