1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Yishan Du, Conrad Borchers, Mutlu Cukurova
As teachers increasingly turn to GenAI in their educational practice, we need robust methods to benchmark large language models (LLMs) for pedagogical purposes. This article presen…