1 paper
Yushi Bai, Jiahao Ying, Yixin Cao +10
Numerous benchmarks have been established to assess the performance of foundation models on open-ended question answering, which serves as a comprehensive test of a model's ability…