5 papers
Searching for Synergy in Shared Workspace Human-AI Collaboration
Nachiket Kotalwar, Rohini Das, Carolyn Rose
Automated AI agents are increasingly capable, yet many scientific and professional tasks require human judgment and contextual expertise. We use simulated shared-workspace human-AI…
Hybrid-Gym: Training Coding Agents to Generalize Across Tasks
Yiqing Xie, Emmy Liu, Gaokai Zhang +7
When assessing the quality of coding agents, predominant benchmarks focus on solving single issues on GitHub, such as SWE-Bench. In contrast, in real use, these agents solve more v…
Inference-Time Personalized Alignment with a Few User Preference Queries
Victor-Alexandru PÄdurean, Parameswaran Kamalaruban, Nachiket Kotalwar +2
We study the problem of aligning a generative model's response with a user's preferences. Recent works have proposed several different formulations for personalized alignment; howe…
Humanizing Automated Programming Feedback: Fine-Tuning Generative Models with Student-Written Feedback
Victor-Alexandru PÄdurean, Tung Phung, Nachiket Kotalwar +4
The growing need for automated and personalized feedback in programming education has led to recent interest in leveraging generative AI for feedback generation. However, current a…
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation
Nachiket Kotalwar, Alkis Gotovos, Adish Singla
Generative AI and large language models hold great promise in enhancing programming education by generating individualized feedback and hints for learners. Recent works have primar…