9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.LG2022
FlexBlock: A Flexible DNN Training Accelerator with Multi-Mode Block Floating Point Support
Seock-Hwan Noh, Jahyun Koo, Seunghyun Lee +2
Training deep neural networks (DNNs) is a computationally expensive job, which can take weeks or months even with high performance GPUs. As a remedy for this challenge, community h…
cs.DC2021★ 9 cited
Multi-model Machine Learning Inference Serving with GPU Spatial Partitioning
Seungbeom Choi, Sunho Lee, Yeonjae Kim +3
As machine learning techniques are applied to a widening range of applications, high throughput machine learning (ML) inference servers have become critical for online service appl…