2 papers
cs.CV2025
Medical Imaging AI Competitions Lack Fairness
Annika Reinke, Evangelia Christodoulou, Sthuthi Sadananda +34
Benchmarking competitions are central to the development of artificial intelligence (AI) in medical imaging, defining performance standards and shaping methodological progress. How…
cs.CV2025
Bridging vision language model (VLM) evaluation gaps with a framework for scalable and cost-effective benchmark generation
Tim Rädsch, Leon Mayer, Simon Pavicic +8
Reliable evaluation of AI models is critical for scientific progress and practical application. While existing VLM benchmarks provide general insights into model capabilities, thei…