Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Emergence WebVoyager: Toward Consistent and Transparent Evaluation of (Web) Agents in The Wild
Deepak Akkil, Mowafak Allaham, Amal Raj +2
Reliable evaluation of AI agents operating in complex, real-world environments requires methodologies that are robust, transparent, and contextually aligned with the tasks agents a…
cs.AI2024
Multimodal Auto Validation For Self-Refinement in Web Agents
Ruhana Azam, Tamer Abuelsaad, Aditya Vempaty +1
As our world digitizes, web agents that can automate complex and monotonous tasks are becoming essential in streamlining workflows. This paper introduces an approach to improving w…
cs.AI2024
Agent-E: From Autonomous Web Navigation to Foundational Design Principles in Agentic Systems
Tamer Abuelsaad, Deepak Akkil, Prasenjit Dey +3
AI Agents are changing the way work gets done, both in consumer and enterprise domains. However, the design patterns and architectures to build highly capable agents or multi-agent…