2 papers
cs.AI2023
MobileNMT: Enabling Translation in 15MB and 30ms
Ye Lin, Xiaohui Wang, Zhexi Zhang +3
Deploying NMT models on mobile devices is essential for privacy, low latency, and offline scenarios. For high model capacity, NMT models are rather large. Running these models on d…
cs.CL2023
Multi-Path Transformer is Better: A Case Study on Neural Machine Translation
Ye Lin, Shuhan Zhou, Yanyang Li +3
For years the model performance in machine learning obeyed a power-law relationship with the model size. For the consideration of parameter efficiency, recent studies focus on incr…