3 papers
cs.AI2026
Benchy: towards a universal language for task-oriented AI benchmarks
Francis F Daniel, Mauro Ibañez, Francis Perelman +1
Benchy is a semantic language and execution engine for benchmarking AI programs. A benchmark is completely specified by a program, a scoring function, and a dataset, B=(P,S,D), and…
cs.AI2026
On the missing benchmarks layer and a potential solution
Francis F Daniel, Mauro Ibañez, Francis Perelman +1
Latin America is missing a foundational layer for native AI development: the benchmark layer. The benchmark layer does two things no other layer can - it audits AI systems against…
cs.AI2026
On the missing data layer and a potential solution
Francis F Daniel, Mauro Ibañez, Francis Perelman +1
Latin America is missing two foundational layers of AI infrastructure: the dataset layer and the benchmark layer. This paper targets the dataset layer. The dataset layer faces two…