4 papers
Beyond Code Generation: LLM-supported Exploration of the Program Design Space
J. D. Zamfirescu-Pereira, Eunice Jun, Michael Terry +2
In this work, we explore explicit Large Language Model (LLM)-powered support for the iterative design of computer programs. Program design, like other design activity, is character…
61A Bot Report: AI Assistants in CS1 Save Students Homework Time and Reduce Demands on Staff. (Now What?)
J. D. Zamfirescu-Pereira, Laryn Qi, Björn Hartmann +2
LLM-based chatbots enable students to get immediate, interactive help on homework assignments, but even a thoughtfully-designed bot may not serve all pedagogical goals. We report h…
A Knowledge-Component-Based Methodology for Evaluating AI Assistants
Laryn Qi, J. D. Zamfirescu-Pereira, Taehan Kim +3
We evaluate an automatic hint generator for CS1 programming assignments powered by GPT-4, a large language model. This system provides natural language guidance about how students…
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
Shreya Shankar, J. D. Zamfirescu-Pereira, Björn Hartmann +2
Due to the cumbersome nature of human evaluation and limitations of code-based evaluation, Large Language Models (LLMs) are increasingly being used to assist humans in evaluating L…