2 papers
cs.LG2026
EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction
Chengxuan Qin, Zhige Chen, Shu Peng +9
Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared task specification layer that…
cs.AI2025
Robustness of Probabilistic Models to Low-Quality Data: A Multi-Perspective Analysis
Liu Peng, Yaochu Jin
A systematic, comparative investigation into the effects of low-quality data reveals a stark spectrum of robustness across modern probabilistic models. We find that autoregressive…