7 papers
Improving Hybrid Human-AI Tutoring by Differentiating Human Tutor Roles Based on Student Needs
Ashish Gurung, Ge Gao, Jordan Gutterman +6
Hybrid human-AI tutoring, where technology and humans jointly facilitate student learning, can be more beneficial than AI-only tutoring. However, preliminary evidence suggests that…
FDARxBench: Benchmarking Regulatory and Clinical Reasoning on FDA Generic Drug Assessment
Betty Xiong, Jillian Fisher, Benjamin Newman +5
We introduce an expert curated, real-world benchmark for evaluating document-grounded question-answering (QA) motivated by generic drug assessment, using the U.S. Food and Drug Adm…
Leveraging LLMs to Assess Tutor Moves in Real-Life Dialogues: A Feasibility Study
Danielle R. Thomas, Conrad Borchers, Jionghao Lin +6
Tutoring improves student achievement, but identifying and studying what tutoring actions are most associated with student learning at scale based on audio transcriptions is an ope…
Detecting LLM-Generated Short Answers and Effects on Learner Performance
Shambhavi Bhushan, Danielle R Thomas, Conrad Borchers +5
The increasing availability of large language models (LLMs) has raised concerns about their potential misuse in online learning. While tools for detecting LLM-generated text exist…
LLM-Generated Feedback Supports Learning If Learners Choose to Use It
Danielle R. Thomas, Conrad Borchers, Shambhavi Bhushan +3
Large language models (LLMs) are increasingly used to generate feedback, yet their impact on learning remains underexplored, especially compared to existing feedback methods. This…
VTutor for High-Impact Tutoring at Scale: Managing Engagement and Real-Time Multi-Screen Monitoring with P2P Connections
Eason Chen, Xinyi Tang, Aprille Xi +5
Hybrid tutoring, where a human tutor supports multiple students in learning with educational technology, is an increasingly common application to deliver high-impact tutoring at sc…