Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
NL2Dashboard: A Lightweight and Controllable Framework for Generating Dashboards with LLMs
Boshen Shi, Kexin Yang, Yuanbo Yang +5
While Large Language Models (LLMs) have demonstrated remarkable proficiency in generating standalone charts, synthesizing comprehensive dashboards remains a formidable challenge. E…
cs.AI2025
MCPWorld: A Unified Benchmarking Testbed for API, GUI, and Hybrid Computer Use Agents
Yunhe Yan, Shihe Wang, Jiajun Du +12
(M)LLM-powered computer use agents (CUA) are emerging as a transformative technique to automate human-computer interaction. However, existing CUA benchmarks predominantly target GU…