111 citations · 420 across the 17 of their papers we have counts for
17 papers
From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL
Wenyue Hua, Zachary Huang, Tyler Payne +3
AI agents increasingly act on their users' behalf, handling tasks such as scheduling meetings, comparing offers, and haggling over prices. These principal-driven tasks routinely pl…
SentinelBench: A Benchmark for Long-Running Monitoring Agents
Matheus Kunzler Maldaner, Adam Fourney, Amanda Swearngin +5
AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: issuing tool calls, refreshin…
Overseeing Agents Without Constant Oversight: Challenges and Opportunities
Madeleine Grunde-McLaughlin, Hussein Mozannar, Maya Murad +3
To enable human oversight, agentic AI systems often provide a trace of reasoning and action steps. Designing traces to have an informative, but not overwhelming, level of detail re…
The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation
Tim R. Davidson, Adam Fourney, Saleema Amershi +3
The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independently developed agents with differen…
Magentic Marketplace: An Open-Source Environment for Studying Agentic Markets
Gagan Bansal, Wenyue Hua, Zezhou Huang +21
As LLM agents advance, they are increasingly mediating economic decisions, ranging from product discovery to transactions, on behalf of users. Such applications promise benefits bu…
Magentic-UI: Towards Human-in-the-loop Agentic Systems
Hussein Mozannar, Gagan Bansal, Cheng Tan +17
AI agents powered by large language models are increasingly capable of autonomously completing complex, multi-step tasks using external tools. Yet, they still fall short of human-l…