3 papers
cs.NI2026
SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference
Hongyao Liu, Liuqun Zhai, Junyi Wang +1
Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the ful…
cs.NI2026
An Efficient Wireless iBCI Headstage with Adaptive ADC Sample Rate
Hongyao Liu, Junyi Wang, Jinglong Chen +1
Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and…
cs.SD2026
ClariCodec: Optimising Neural Speech Codes for 200bps Communication using Reinforcement Learning
Junyi Wang, Chi Zhang, Jing Qian +4
In bandwidth-constrained communication such as satellite and underwater channels, speech must often be transmitted at ultra-low bitrates where intelligibility is the primary object…