Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Medical Image Spatial Grounding with Semantic Sampling
Andrew Seohwan Yu, Mohsen Hariri, Kunio Nakamura +3
Vision language models (VLMs) have shown significant promise in visual grounding for images as well as videos. In medical imaging research, VLMs represent a bridge between object d…
cs.CV2024
Novel adaptation of video segmentation to 3D MRI: efficient zero-shot knee segmentation with SAM2
Andrew Seohwan Yu, Mohsen Hariri, Xuecen Zhang +3
Intelligent medical image segmentation methods are rapidly evolving and being increasingly applied, yet they face the challenge of domain transfer, where algorithm performance degr…