1 citations · 1 across the 5 of their papers we have counts for
1 paper · 1 filter
Stefan Broecker, Mason del Rosario, Boris Selitser +1
The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the training and evaluation of these mod…