From the 1 of 4 linked papers with an AI index.
4 papers
PTQ4SNN: Membrane-Aware Post-Training Quantization for Spiking Neural Networks
Hui Xie, Tong Shi, Haotong Qin +3
Spiking neural networks (SNNs) enable sparse and event-driven computation, but their low-bit deployment remains incomplete because recurrent membrane states are commonly retained i…
SemPIC: Learning Semantic Position-Independent KV Caches
Hui Xie, Peng Xiao, Yutong Deng +4
The paper introduces SemPIC, a method that learns semantic position‑independent key‑value caches for large language models by training a LoRA‑enabled writer to compile document rep…
An Empirical Study of openPangu Quantization on Ascend NPUs
Tong Shi, Jiacheng Wang, Hui Xie +4
openPangu models are attractive targets for private and domestic large-language-model deployment, yet their robustness under aggressive post-training quantization on Ascend NPUs ha…
SPEAR: Structured Pruning for Spiking Neural Networks via Synaptic Operation Estimation and Reinforcement Learning
Hui Xie, Yuhe Liu, Shaoqi Yang +6
While deep spiking neural networks (SNNs) demonstrate superior performance, their deployment on resource-constrained neuromorphic hardware still remains challenging. Network prunin…