3 papers
cs.LG2024
Research on Key Technologies for Cross-Cloud Federated Training of Large Language Models
Haowei Yang, Mingxiu Sui, Shaobo Liu +3
With the rapid development of natural language processing technology, large language models have demonstrated exceptional performance in various application scenarios. However, tra…
cs.CL2024
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
Zhen Qi, Jiajing Chen, Shuo Wang +3
This study aims to explore the performance improvement method of large language models based on GPT-4 under the multi-task learning framework and conducts experiments on two tasks:…
cs.AI2024
Adaptive Optimization for Enhanced Efficiency in Large-Scale Language Model Training
Jiajing Chen, Bingying Liu, Xiaoxuan Liao +3
With the rapid development of natural language processing technology, large-scale language models (LLM) have achieved remarkable results in a variety of tasks. However, how to effe…