1 paper
Sajjad Kachuee, Mohammad Sharifkhani
Adjusting the latency, power, and accuracy of natural language understanding models is a desirable objective of an efficient architecture. This paper proposes an efficient Transfor…