1 citations · 1 across the 1 of their papers we have counts for
1 paper
Luoxi Tang, Tharunya Sundar, Yuqiao Meng +7
As large language models (LLMs) are increasingly integrated into educational tools, current evaluations on standardized tests predominantly focus on binary outcome accuracy. Instea…