5 papers
Scientific judgment drifts over time in AI ideation
Lingyu Zhang, Mitchell Wang, Boyuan Chen
Scientific discovery begins with ideas, yet evaluating early-stage research concepts is a subtle and subjective human judgment. As large language models (LLMs) are increasingly tas…
Pref-GUIDE: Continual Policy Learning from Real-Time Human Feedback via Preference-Based Learning
Zhengran Ji, Boyuan Chen
Training reinforcement learning agents with human feedback is crucial when task objectives are difficult to specify through dense reward functions. While prior methods rely on offl…
CREW-WILDFIRE: Benchmarking Agentic Multi-Agent Collaborations at Scale
Jonathan Hyun, Nicholas R Waytowich, Boyuan Chen
Despite rapid progress in large language model (LLM)-based multi-agent systems, current benchmarks fall short in evaluating their scalability, robustness, and coordination capabili…
GUIDE: Real-Time Human-Shaped Agents
Lingyu Zhang, Zhengran Ji, Nicholas R Waytowich +1
The recent rapid advancement of machine learning has been driven by increasingly powerful models with the growing availability of training data and computational resources. However…
Enabling Multi-Robot Collaboration from Single-Human Guidance
Zhengran Ji, Lingyu Zhang, Paul Sajda +1
Learning collaborative behaviors is essential for multi-agent systems. Traditionally, multi-agent reinforcement learning solves this implicitly through a joint reward and centraliz…