1 paper
Sangjun Lee, Seung-taek Woo, Jungyu Jin +2
To enable broader deployment of Large Language Models (LLMs), it is essential to identify the best-performing model under strict memory constraints. We present AMQ, Automated Mixed…