3 papers
cs.MA2026
Training Small LLMs as Spatial Multi-Agent Policies
Yi Mao, Andrew Perrault
Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that such systems should be judged by…
cs.LG2026
WARC-Bench: Web Archive Based Benchmark for GUI Subtask Executions
Sanjari Srivastava, Gang Li, Cheng Chang +8
Training web agents to navigate complex, real-world websites requires them to master - short-horizon interactions on multiple UI components (e.g., choosing the…
cs.LG2025
Optimizing Urban Service Allocation with Time-Constrained Restless Bandits
Yi Mao, Andrew Perrault
Municipal inspections are an important part of maintaining the quality of goods and services. In this paper, we approach the problem of intelligently scheduling service inspections…