3 papers
cs.CV2026
ESICA: A Scalable Framework for Text-Guided 3D Medical Image Segmentation
Yu Xin, Gorkem Can Ates, Jun Ma +5
Text guided 3D medical image segmentation offers a flexible alternative to class based and spatial prompt based models by allowing users to specify regions of interest directly in…
cs.CV2025
Med3DVLM: An Efficient Vision-Language Model for 3D Medical Image Analysis
Yu Xin, Gorkem Can Ates, Kuang Gong +1
Vision-language models (VLMs) have shown promise in 2D medical image analysis, but extending them to 3D remains challenging due to the high computational demands of volumetric data…
cs.CV2025
DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions
Gorkem Can Ates, Yu Xin, Kuang Gong +1
Vision-language models (VLMs) have been widely applied to 2D medical image analysis due to their ability to align visual and textual representations. However, extending VLMs to 3D…