2 citations · 3 across the 5 of their papers we have counts for
Showing cs.MAShow all
2 papers · 1 filter
cs.MA2026
Emergence World: Adversarial Stress-Testing of Long-Horizon Multi-Agent Systems
Deepak Akkil, Tamer Abuelsaad, Karthik Vikram +5
As AI agents move from bounded tasks to persistent deployments, failures can propagate through memory, tools, other agents, and environmental state long after their interactions. T…
cs.MA2026
Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy
Deepak Akkil, Ravi Kokku, Karthik Vikram +3
Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mismatched with the deployment con…