2 papers
cs.AI2026
When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems
Zehao Wang, Shilong Jin, Zhao Cao +1
LLM-based multi-agent systems can fail even when planned actions are executed correctly because agents may misjudge their knowledge when evaluating plan feasibility, a phenomenon w…
cs.LG2026
SE-GA: Memory-Augmented Self-Evolution for GUI Agents
Shilong Jin, Lanjun Wang, Zhuosheng Zhang
Autonomous Graphical User Interface (GUI) agents often struggle with multi-step tasks due to constrained context windows and static policies that fail to adapt to dynamic environme…