1 paper · 1 filter
Charles Eckert, Xiaowei Wang, Jingcheng Wang +5
This paper presents the Neural Cache architecture, which re-purposes cache structures to transform them into massively parallel compute units capable of running inferences for Deep…