1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Lequn Chen, Weixin Deng, Anirudh Canumalla +4
Having large batch sizes is one of the most critical aspects of increasing the accelerator efficiency and the performance of DNN model inference. However, existing model serving sy…