2 papers
cs.LG2025
CaTE Data Curation for Trustworthy AI
Mary Versa Clemens-Sewall, Christopher Cervantes, Emma Rafkin +6
This report provides practical guidance to teams designing or developing AI-enabled systems for how to promote trustworthiness during the data curation phase of development. In thi…
physics.data-an2025
Towards a Large Physics Benchmark
Kristian G. Barman, Sascha Caron, Faegheh Hasibi +5
We introduce a benchmark framework developed by and for the scientific community to evaluate, monitor and steer large language model development in fundamental physics. Building on…