1 paper · 1 filter
Minsoo Kim, Kyuhong Shim, Seongmin Park +2
Pre-trained Transformer models such as BERT have shown great success in a wide range of applications, but at the cost of substantial increases in model complexity. Quantization-awa…