2 papers
cs.NI2026
SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference
Hongyao Liu, Liuqun Zhai, Junyi Wang +1
Efficient inference for on-device Large Language Models (LLMs) remains challenging due to limited hardware resources and the high cost of the prefill stage, which processes the ful…
cs.NI2026
An Efficient Wireless iBCI Headstage with Adaptive ADC Sample Rate
Hongyao Liu, Junyi Wang, Jinglong Chen +1
Implantable Brain-Computer Interfaces (iBCIs) are increasingly pivotal in clinical and daily applications. However, wireless iBCIs face severe constraints in power consumption and…