Showing cs.NEShow all
3 papers · 1 filter
cs.NE2025
Spatio-Temporal Pruning for Compressed Spiking Large Language Models
Yi Jiang, Malyaban Bal, Brian Matejek +3
Large Language Models (LLMs) present significant challenges for deployment in energy-constrained environments due to their large model sizes and high inference latency. Spiking Neu…
cs.NE2024
Exploring Extreme Quantization in Spiking Language Models
Malyaban Bal, Yi Jiang, Abhronil Sengupta
Despite the growing prevalence of large language model (LLM) architectures, a crucial concern persists regarding their energy and power consumption, which still lags far behind the…
cs.NE2024
Stochastic Spiking Neural Networks with First-to-Spike Coding
Yi Jiang, Sen Lu, Abhronil Sengupta
Spiking Neural Networks (SNNs), recognized as the third generation of neural networks, are known for their bio-plausibility and energy efficiency, especially when implemented on ne…