2 papers
cs.CV2026
HeiCo-FOCUS: A Clinically Grounded Dataset for Long-Context Video Understanding
Leon Mayer, Lucas Luttner, Patrick Godau +42
Recent advances in Vision-Language Models (VLMs) have led to rapid progress in video understanding across a wide range of benchmark tasks. However, existing evaluations largely foc…
cs.CV2025
XiCAD: Camera Activation Detection in the Da Vinci Xi User Interface
Alexander C. Jenke, Gregor Just, Claas de Boer +3
Purpose: Robot-assisted minimally invasive surgery relies on endoscopic video as the sole intraoperative visual feedback. The DaVinci Xi system overlays a graphical user interface…