1 paper · 1 filter
Ayan Sengupta, Siddhant Chaudhary, Tanmoy Chakraborty
Scaling up model parameters and training data consistently improves the performance of large language models (LLMs), but at the cost of rapidly growing memory and compute requireme…