2 citations · 2 across the 1 of their papers we have counts for
1 paper
Maxime Lelièvre, Amy Waldock, Meng Liu +7
Benchmarks like Massive Multitask Language Understanding (MMLU) have played a pivotal role in evaluating AI's knowledge and abilities across diverse domains. However, existing benc…