5 citations · 8 across the 5 of their papers we have counts for
4 papers · 1 filter
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
Shuo Wang, Chihang Wang, Jia Gao +3
This study proposes a knowledge distillation algorithm based on large language models and feature alignment, aiming to effectively transfer the knowledge of large pre-trained model…
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks
Jiacheng Hu, Xiaoxuan Liao, Jia Gao +3
This study proposes a large language model optimization method based on the improved LoRA fine-tuning algorithm, aiming to improve the accuracy and computational efficiency of the…
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
Zhen Qi, Jiajing Chen, Shuo Wang +3
This study aims to explore the performance improvement method of large language models based on GPT-4 under the multi-task learning framework and conducts experiments on two tasks:…
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation
Jiajing Chen, Shuo Wang, Zhen Qi +3
This research introduces a novel text generation model that combines BERT's semantic interpretation strengths with GPT-4's generative capabilities, establishing a high standard in…