15 citations · 20 across the 7 of their papers we have counts for
4 papers · 1 filter
Decouple and Orthogonalize: A Data-Free Framework for LoRA Merging
Shenghe Zheng, Hongzhi Wang, Chenyu Huang +5
With more open-source models available for diverse tasks, model merging has gained attention by combining models into one, reducing training, storage, and inference costs. Current…
Dynamic Base model Shift for Delta Compression
Chenyu Huang, Peng Ye, Shenghe Zheng +4
Transformer-based models with the pretrain-finetune paradigm bring about significant progress, along with the heavy storage and deployment costs of finetuned models on multiple tas…
Seeing Delta Parameters as JPEG Images: Data-Free Delta Compression with Discrete Cosine Transform
Chenyu Huang, Peng Ye, Xiaohui Wang +5
With transformer-based models and the pretrain-finetune paradigm becoming mainstream, the high storage and deployment costs of individual finetuned models on multiple tasks pose cr…
Stimulative Training of Residual Networks: A Social Psychology Perspective of Loafing
Peng Ye, Shengji Tang, Baopu Li +2
Residual networks have shown great success and become indispensable in today's deep models. In this work, we aim to re-investigate the training process of residual networks from a…