4 citations · 6 across the 4 of their papers we have counts for
3 papers · 1 filter
Assessing the Quality of AI-Generated Exams: A Large-Scale Field Study
Calvin Isley, Joshua Gilbert, Evangelos Kassos +9
While large language models (LLMs) challenge conventional methods of teaching and learning, they present an exciting opportunity to improve efficiency and scale high-quality instru…
Measurement to Meaning: A Validity-Centered Framework for AI Evaluation
Olawale Salaudeen, Anka Reuel, Ahmed Ahmed +6
While the capabilities and utility of AI systems have advanced, rigorous norms for evaluating these systems have lagged. Grand claims, such as models achieving general reasoning ca…
AI and Holistic Review: Informing Human Reading in College Admissions
AJ Alvero, Noah Arthurs, anthony lising antonio +4
College admissions in the United States is carried out by a human-centered method of evaluation known as holistic review, which typically involves reading original narrative essays…