4 citations · 4 across the 1 of their papers we have counts for
2 papers
cs.CV2026
HeiCo-FOCUS: A Clinically Grounded Dataset for Long-Context Video Understanding
Leon Mayer, Lucas Luttner, Patrick Godau +42
Recent advances in Vision-Language Models (VLMs) have led to rapid progress in video understanding across a wide range of benchmark tasks. However, existing evaluations largely foc…
cs.CV2025★ 4 cited
Comparative validation of surgical phase recognition, instrument keypoint estimation, and instrument instance segmentation in endoscopy: Results of the PhaKIR 2024 challenge
Tobias Rueckert, David Rauber, Raphaela Maerkl +58
Reliable recognition and localization of surgical instruments in endoscopic video recordings are foundational for a wide range of applications in computer- and robot-assisted minim…