3 papers
cs.AI2026
Build, Judge, Optimize: A Blueprint for Continuous Improvement of Multi-Agent Consumer Assistants
Alejandro Breen Herrera, Aayush Sheth, Steven G. Xu +8
Conversational shopping assistants (CSAs) represent a compelling application of agentic AI, but moving from prototype to production reveals two underexplored challenges: how to eva…
cs.AI2026
Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work
Zihao Wu, Steven Xu, Bowen Chen +8
With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people code, making it more accessibl…
cs.CL2025
Large language models management of medications: three performance analyses
Kelli Henry, Steven Xu, Kaitlin Blotske +11
Purpose: Large language models (LLMs) have proven performance for certain diagnostic tasks, however limited studies have evaluated their consistency in recommending appropriate med…