From the 1 of 2 linked papers with an AI index.
2 papers
cs.CV2026
Invisible Ink Threats: Adversarial Goals Behind Legitimate Tasks in Computer-Use Agents
Jia-Chen Zhang, Ze-Yu Zhang, Kai-Wei Zhang
Computer-use agents (CUAs), which empower large language models to autonomously operate operating systems and the web, are increasingly vulnerable to indirect prompt injection atta…
cs.AI2026
OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding
Jingbo Zhou, Yusai Zhao, Qi Bao +12
The paper presents OmegaUse-OfficeVal, a benchmark that evaluates large language model agents on long‑horizon office‑suite tasks while providing economic signals (human labor time…