Codabench: Flexible, Easy-to-Use and Reproducible Benchmarking Platform
arXiv:2110.05802 · doi:10.1016/j.patter.2022.100543
Abstract
Obtaining standardized crowdsourced benchmark of computational methods is a major issue in data science communities. Dedicated frameworks enabling fair benchmarking in a unified environment are yet to be developed. Here we introduce Codabench, an open-source, community-driven platform for benchmarking algorithms or software agents versus datasets or tasks. A public instance of Codabench (https://www.codabench.org/) is open to everyone, free of charge, and allows benchmark organizers to compare fairly submissions, under the same setting (software, hardware, data, algorithms), with custom protocols and data formats. Codabench has unique features facilitating the organization of benchmarks flexibly, easily and reproducibly, such as the possibility of re-using templates of benchmarks, and supplying compute resources on-demand. Codabench has been used internally and externally on various applications, receiving more than 130 users and 2500 submissions. As illustrative use cases, we introduce 4 diverse benchmarks covering Graph Machine Learning, Cancer Heterogeneity, Clinical Diagnosis and Reinforcement Learning.
References in corpus (1)
Cited by in corpus (5)
- ICPR 2024 Competition on Domain Adaptation and GEneralization for Character Classification (DAGECC)
- Data Publishing in Mechanics and Dynamics: Challenges, Guidelines, and Examples from Engineering Design
- First International StepUP Competition for Biometric Footstep Recognition: Methods, Results and Remaining Challenges
- AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report
- AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance