2 papers
cs.DC2026
LMDeploy Accelerates Mixed-Precision LLM Inference with TurboMind
Li Zhang, Youhe Jiang, Guoliang He +6
Mixed-precision inference techniques reduce the memory and computational demands of Large Language Models (LLMs) by applying hybrid precision formats to model weights, activations,…
cs.HC2024
DuA: Dual Attentive Transformer in Long-Term Continuous EEG Emotion Analysis
Yue Pan, Qile Liu, Qing Liu +6
Affective brain-computer interfaces (aBCIs) are increasingly recognized for their potential in monitoring and interpreting emotional states through electroencephalography (EEG) sig…