4 papers
HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents
Chao Jin, Wenkui Yang, Hao Sun +6
While progress in GUI agents has been largely driven by industrial-scale training, ungrounded hallucinations often trigger cascading failures in real-world deployments.Unlike gener…
Training-Free Test-Time Contrastive Learning for Large Language Models
Kaiwen Zheng, Kai Zhou, Jinwu Hu +3
Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-time adaptation (TTA) methods r…
EE-MCP: Self-Evolving MCP-GUI Agents via Automated Environment Generation and Experience Learning
Tiantian He, Yihang Chen, Keyue Jiang +4
Computer-use agents that combine GUI interaction with structured API calls via the Model Context Protocol (MCP) show promise for automating software tasks. However, existing approa…
Automating Agentic Workflow Generation via Self-Adaptive Abstraction Operators
Mingming Zhao, Xiaokang Wei, Yuanqi Shao +5
Large language models (LLMs) have shown strong potential in automating the design of agentic workflows. However, existing methods still rely heavily on manually predefined operator…