1 citations · 1 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
SurgCheck: Do Vision-Language Models Really Look at Images in Surgical VQA?
Jongmin Shin, Ka Young Kim, Eunki Cho +2
Purpose: Vision-language models (VLMs) have shown promising performance in surgical visual question answering (VQA). However, existing surgical VQA datasets often contain linguisti…
cs.CV2026
CurConMix+: A Unified Spatio-Temporal Framework for Hierarchical Surgical Workflow Understanding
Yongjun Jeon, Jongmin Shin, Kanggil Park +8
Surgical action triplet recognition aims to understand fine-grained surgical behaviors by modeling the interactions among instruments, actions, and anatomical targets. Despite its…