'A bit of chaos and madness': The AI Assessment Scale and the work of assessment reform
arXiv:2606.26729 · doi:10.1080/02602938.2026.2724300
Abstract
Generative artificial intelligence (GenAI) has intensified pressure on universities to redesign assessment while maintaining integrity, equity, and validity. Structured frameworks such as the Artificial Intelligence Assessment Scale (AIAS) offer one response, but evidence of how staff experience their implementation remains limited. This qualitative study examines AIAS implementation at a private international university in Vietnam and a public university in the United Kingdom. Data from five focus groups with 30 academic staff were analysed using hybrid thematic analysis, with Critical AI Literacy used as a sensitising concept. Six themes were developed: recognising and integrating AI, facilitating conditions, building capacity, pathways to adoption, ethics in practice, and reframing pedagogy. Staff valued the AIAS as a shared language for legitimising GenAI use, clarifying boundaries, and prompting reflection on assessment design. However, implementation was shaped by governance, tool access, staff confidence, workload, integrity concerns, disciplinary context, and alignment with learning outcomes. The findings show that the AIAS could prompt authentic assessment design and student engagement, but may become a compliance layer when disconnected from learning outcomes, disciplinary context, and staff capacity. This study contributes empirical evidence on the institutional conditions through which GenAI assessment frameworks move from policy adoption to pedagogical enactment.
V2: Corrections of errata, additional limitations
References in corpus (5)
- The AI Assessment Scale (AIAS): A Framework for Ethical Integration of Generative AI in Educational Assessment
- The AI Assessment Scale Revisited: A Framework for Educational Assessment
- Funhouse Mirror or Echo Chamber? A Methodological Approach to Teaching Critical AI Literacy Through Metaphors
- Assessment Twins: A Protocol for AI-Vulnerable Summative Assessment
- Dramaturgies of Deception: AI Humanizers and the Performance of Legitimacy in Higher Education Assessment