2 papers
cs.HC2026
Calibrating Trustworthiness: Co-Designing Metrics and Visualizations for Evaluating LLMs in Education
Adam Coscia, Sujata Duwal, Langdon Holmes +2
LLMs are reshaping educational technology, yet evaluating their responses for pedagogical alignment remains underexplored, relying heavily on the expertise of learning engineers bu…
cs.AI2026
LLM Prompt Evaluation for Educational Applications
Langdon Holmes, Adam Coscia, Scott Crossley +2
As large language models (LLMs) become increasingly common in educational applications, there is a growing need for evidence-based methods to design and evaluate LLM prompts that p…