1 citations · 1 across the 11 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Stop Comparing LLM Agents Without Disclosing the Harness
Yunbei Zhang, Janet Wang, Yingqiang Ge +3
This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely the infrastructure layer th…
cs.AI2026
ClawSafety: "Safe" LLMs, Unsafe Agents
Bowen Wei, Yunbei Zhang, Jinhao Pan +5
Personal AI agents like OpenClaw run with elevated privileges on users' local machines, where a single successful prompt injection can leak credentials, redirect financial transact…
cs.AI2025
eSkinHealth: A Multimodal Dataset for Neglected Tropical Skin Diseases
Janet Wang, Xin Hu, Yunbei Zhang +7
Skin Neglected Tropical Diseases (NTDs) impose severe health and socioeconomic burdens in impoverished tropical communities. Yet, advancements in AI-driven diagnostic support are h…