2 papers
cs.CL2026
PICon: A Multi-Turn Interrogation Framework for Evaluating Persona Agent Consistency
Minseo Kim, Sujeong Im, Junseong Choi +4
Large language model (LLM)-based persona agents are rapidly being adopted as scalable proxies for human participants across diverse domains. Yet there is no systematic method for v…
cs.CY2025
A Large-Scale Real-World Evaluation of LLM-Based Virtual Teaching Assistant
Sunjun Kweon, Sooyohn Nam, Hyunseung Lim +2
Virtual Teaching Assistants (VTAs) powered by Large Language Models (LLMs) have the potential to enhance student learning by providing instant feedback and facilitating multi-turn…