2 papers
cs.AI2026
Unsteady Metrics and Benchmarking Cultures of AI Model Builders
Stefan Baack, Christo Buschek, Maty Bohacek
The primary way to establish and compare competencies in foundation and generative AI models has shifted from peer-reviewed literature to press releases and company blog posts, whe…
cs.CY2025
Towards Best Practices for Open Datasets for LLM Training
Stefan Baack, Stella Biderman, Kasia Odrozek +36
Many AI companies are training their large language models (LLMs) on data without the permission of the copyright owners. The permissibility of doing so varies by jurisdiction: in…