2 papers
cs.AR2019
Near Data Acceleration with Concurrent Host Access
Benjamin Y. Cho, Yongkee Kwon, Sangkug Lym +1
Near-data accelerators (NDAs) that are integrated with main memory have the potential for significant power and performance benefits. Fully realizing these benefits requires the la…
cs.LG2018
Mini-batch Serialization: CNN Training with Inter-layer Data Reuse
Sangkug Lym, Armand Behroozi, Wei Wen +3
Training convolutional neural networks (CNNs) requires intense computations and high memory bandwidth. We find that bandwidth today is over-provisioned because most memory accesses…