2 papers
cs.AR2026
Salca: A Sparsity-Aware Hardware Accelerator for Efficient Long-Context Attention Decoding
Wang Fan, Wei Cao, Xi Zha +5
Long contexts improve capabilities of large language models but pose serious hardware challenges: compute and memory footprints grow linearly with sequence length. Particularly, th…
cs.LG2025
SpikeSTAG: Spatial-Temporal Forecasting via GNN-SNN Collaboration
Bang Hu, Changze Lv, Mingjie Li +5
Spiking neural networks (SNNs), inspired by the spiking behavior of biological neurons, offer a distinctive approach for capturing the complexities of temporal data. However, their…