collaborators

6 papers

cs.CL2026

Hallucination Detection and Evaluation of Large Language Model

Chenggong Zhang, Haopeng Wang, Hexi Meng

Hallucinations in Large Language Models (LLMs) pose a significant challenge, generating misleading or unverifiable content that undermines trust and reliability. Existing evaluatio…

cs.AI2026

MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion

Yunfei Feng, Xi Zhao, Cheng Zhang +5

Mobile agents can autonomously complete user-assigned tasks through GUI interactions. However, existing mainstream evaluation benchmarks, such as AndroidWorld, operate by connectin…

cs.AI2025

Web World Models

Jichen Feng, Yifan Zhang, Chenggong Zhang +3

Language agents increasingly require persistent worlds in which they can act, remember, and learn. Existing approaches sit at two extremes: conventional web frameworks provide reli…

cs.AI2025

Beyond Training: Enabling Self-Evolution of Agents with MOBIMEM

Zibin Liu, Cheng Zhang, Xi Zhao +6

Large Language Model (LLM) agents are increasingly deployed to automate complex workflows in mobile and desktop environments. However, current model-centric agent architectures str…

cs.MA2025

MobiAgent: A Systematic Framework for Customizable Mobile Agents

Cheng Zhang, Erhu Feng, Xi Zhao +7

With the rapid advancement of Vision-Language Models (VLMs), GUI-based mobile agents have emerged as a key development direction for intelligent mobile systems. However, existing a…

cs.LG2025

Get Experience from Practice: LLM Agents with Record & Replay

Erhu Feng, Wenbo Zhou, Zibin Liu +8

AI agents, empowered by Large Language Models (LLMs) and communication protocols such as MCP and A2A, have rapidly evolved from simple chatbots to autonomous entities capable of ex…