1 paper · 1 filter
Ganghua Wang, Zhaorun Chen, Bo Li +1
As foundation models continue to scale, the size of trained models grows exponentially, presenting significant challenges for their evaluation. Current evaluation practices involve…