2 papers
cs.AI2026
OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding
Jingbo Zhou, Yusai Zhao, Qi Bao +12
Large language model (LLM) agents are increasingly expected to assist users in completing tasks. However, existing benchmarks provide limited support for evaluating whether agents…
cs.AI2026
ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?
Lingwei Li, Yirong Kan, Peng Chen +3
This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and locally deployable without sacrificing accuracy. EEG foundation models ar…