7 papers
MultEval: Supporting Collaborative Alignment for LLM-as-a-Judge Evaluation Criteria
Charles Chiang, Simret Gebreegziabher, Annalisa Szymanski +6
LLM-as-a-judge approaches have emerged as a scalable solution for evaluating model behaviors, yet they rely on evaluation criteria often created by a single individual, embedding t…
From Verification Burden to Trusted Collaboration: Design Goals for LLM-Assisted Literature Reviews
Brenda Nogueira, Werner Geyer, Andrew Anderson +4
Large Language Models (LLMs) are increasingly embedded in academic writing practices. Although numerous studies have explored how researchers employ these tools for scientific writ…
ALLOY: Generating Reusable Agent Workflows from User Demonstration
Jiawen Li, Zheng Ning, Yuan Tian +1
Large language models (LLMs) enable end-users to delegate complex tasks to autonomous agents through natural language. However, prompt-based interaction faces critical limitations:…
AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multi-modal Information Between Reality and Videos
Zheng Ning, Leyang Li, Daniel Killough +6
Videos offer rich audiovisual information that can support people in performing activities of daily living (ADLs), but they remain largely inaccessible to blind or low-vision (BLV)…
GLITTER: An AI-assisted Platform for Material-Grounded Asynchronous Discussion in Flipped Learning
Weirui Peng, Yinuo Yang, Zheng Zhang +1
Flipped classrooms promote active learning by having students engage with materials independently before class, allowing in-class time for collaborative problem-solving. During thi…
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
Yuan Tian, Jonathan K. Kummerfeld, Toby Jia-Jun Li +1
Though recent advances in machine learning have led to significant improvements in natural language interfaces for databases, the accuracy and reliability of these systems remain l…