3 papers
cs.CL2024
Gated Slot Attention for Efficient Linear-Time Sequence Modeling
Yu Zhang, Songlin Yang, Ruijie Zhu +9
Linear attention Transformers and their gated variants, celebrated for enabling parallel training and efficient recurrent inference, still fall short in recall-intensive tasks comp…
eess.SP2024
A Survey of Spatio-Temporal EEG data Analysis: from Models to Applications
Pengfei Wang, Huanran Zheng, Silong Dai +4
In recent years, the field of electroencephalography (EEG) analysis has witnessed remarkable advancements, driven by the integration of machine learning and artificial intelligence…
cs.CL2024
Scalable MatMul-free Language Modeling
Rui-Jie Zhu, Yu Zhang, Steven Abreu +7
Large Language Models (LLMs) have fundamentally altered how we approach scaling in machine learning. However, these models pose substantial computational and memory challenges, pri…