1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2026
HeiCo-FOCUS: A Clinically Grounded Dataset for Long-Context Video Understanding
Leon Mayer, Lucas Luttner, Patrick Godau +42
Recent advances in Vision-Language Models (VLMs) have led to rapid progress in video understanding across a wide range of benchmark tasks. However, existing evaluations largely foc…
cs.CV2025
Challenging Vision-Language Models with Surgical Data: A New Dataset and Broad Benchmarking Study
Leon Mayer, Tim Rädsch, Dominik Michael +8
While traditional computer vision models have historically struggled to generalize to endoscopic domains, the emergence of foundation models has shown promising cross-domain perfor…
cs.CV2023★ 1 cited
Self-distillation for surgical action recognition
Amine Yamlahi, Thuy Nuong Tran, Patrick Godau +9
Surgical scene understanding is a key prerequisite for contextaware decision support in the operating room. While deep learning-based approaches have already reached or even surpas…