2 papers
cs.LG2021
Low-Precision Hardware Architectures Meet Recommendation Model Inference at Scale
Zhaoxia, Deng, Jongsoo Park +17
Tremendous success of machine learning (ML) and the unabated growth in ML model complexity motivated many ML-specific designs in both CPU and accelerator architectures to speed up…
cs.LG2020
Mixed-Precision Embedding Using a Cache
Jie Amy Yang, Jianyu Huang, Jongsoo Park +2
In recommendation systems, practitioners observed that increase in the number of embedding tables and their sizes often leads to significant improvement in model performances. Give…