3 papers
cs.LG2025
Graph-Based Spectral Decomposition for Parameter Coordination in Language Model Fine-Tuning
Hanlu Zhang, Yumeng Ma, Shuo Wang +2
This paper proposes a parameter collaborative optimization algorithm for large language models, enhanced with graph spectral analysis. The goal is to improve both fine-tuning effic…
cs.CL2024
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
Shuo Wang, Chihang Wang, Jia Gao +3
This study proposes a knowledge distillation algorithm based on large language models and feature alignment, aiming to effectively transfer the knowledge of large pre-trained model…
cs.CL2024
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
Zhen Qi, Jiajing Chen, Shuo Wang +3
This study aims to explore the performance improvement method of large language models based on GPT-4 under the multi-task learning framework and conducts experiments on two tasks:…