3 papers
stat.ML2024
BanditCAT and AutoIRT: Machine Learning Approaches to Computerized Adaptive Testing and Item Calibration
James Sharpnack, Kevin Hao, Phoebe Mulcaire +4
In this paper, we present a complete framework for quickly calibrating and administering a robust large-scale computerized adaptive test (CAT) with a small number of responses. Cal…
cs.LG2024
AutoIRT: Calibrating Item Response Theory Models with Automated Machine Learning
James Sharpnack, Phoebe Mulcaire, Klinton Bicknell +2
Item response theory (IRT) is a class of interpretable factor models that are widely used in computerized adaptive tests (CATs), such as language proficiency tests. Traditionally,…
cs.CY2024
Responsible AI for Test Equity and Quality: The Duolingo English Test as a Case Study
Jill Burstein, Geoffrey T. LaFlair, Kevin Yancey +2
Artificial intelligence (AI) creates opportunities for assessments, such as efficiencies for item generation and scoring of spoken and written responses. At the same time, it poses…