1 paper
Elise Paradis, Ambar Murillo, Maulishree Pandey +4
In the AI community, benchmarks to evaluate model quality are well established, but an equivalent approach to benchmarking products built upon generative AI models is still missing…