3 papers
cs.CL2025
Completion Collaboration: Scaling Collaborative Effort with Agents
Shannon Zejiang Shen, Valerie Chen, Ken Gu +11
Current evaluations of agents remain centered around one-shot task completion, failing to account for the inherently iterative and collaborative nature of many real-world problems,…
cs.CL2025
Diagnosing our datasets: How does my language model learn clinical information?
Furong Jia, David Sontag, Monica Agrawal
Large language models (LLMs) have performed well across various clinical natural language processing tasks, despite not being directly trained on electronic health record (EHR) dat…
cs.HC2025
CodingGenie: A Proactive LLM-Powered Programming Assistant
Sebastian Zhao, Alan Zhu, Hussein Mozannar +3
While developers increasingly adopt tools powered by large language models (LLMs) in day-to-day workflows, these tools still require explicit user invocation. To seamlessly integra…