1 paper · 1 filter
Sungbin Shin, Wonpyo Park, Jaeho Lee +1
This work suggests fundamentally rethinking the current practice of pruning large language models (LLMs). The way it is done is by divide and conquer: split the model into submodel…