Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Touchstone Benchmark: Are We on the Right Way for Evaluating AI Algorithms for Medical Segmentation?
Pedro R. A. S. Bassi, Wenxuan Li, Yucheng Tang +50
How can we test AI performance? This question seems trivial, but it isn't. Standard benchmarks often have problems such as in-distribution and small-size test sets, oversimplified…
cs.CV2024
Decoupling Semantic Similarity from Spatial Alignment for Neural Networks
Tassilo Wald, Constantin Ulrich, Gregor Köhler +6
What representation do deep neural networks learn? How similar are images to each other for neural networks? Despite the overwhelming success of deep learning methods key questions…