3 papers
cs.LG2026
EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction
Chengxuan Qin, Zhige Chen, Shu Peng +9
Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared task specification layer that…
cs.LG2026
ProCUA-SFT Technical Report
Jaehun Jung, Ximing Lu, Brandon Cui +11
Training computer-use agents (CUAs) -- models that interact with graphical desktops through screenshots and keyboard/mouse actions -- requires large-scale, diverse trajectory data…
cs.SE2024
MRWeb: An Exploration of Generating Multi-Page Resource-Aware Web Code from UI Designs
Yuxuan Wan, Yi Dong, Jingyu Xiao +3
Multi-page websites dominate modern web development. However, existing design-to-code methods rely on simplified assumptions, limiting to single-page, self-contained webpages witho…