1 citations · 1 across the 1 of their papers we have counts for
3 papers · 1 filter
SAP-Bench: Benchmarking Multimodal Large Language Models in Surgical Action Planning
Mengya Xu, Zhongzhen Huang, Dillan Imans +3
Effective evaluation is critical for driving advancements in MLLM research. The surgical action planning (SAP) task, which aims to generate future action sequences from visual inpu…
SurgVLM: A Large Vision-Language Model and Systematic Evaluation Benchmark for Surgical Intelligence
Zhitao Zeng, Zhu Zhuo, Xiaojun Jia +12
Foundation models have achieved transformative success across biomedical domains by enabling holistic understanding of multimodal data. However, their application in surgery remain…
Why is the winner the best?
Matthias Eisenmann, Annika Reinke, Vivienn Weru +122
International benchmarking competitions have become fundamental for the comparative performance assessment of image analysis methods. However, little attention has been given to in…