6 papers
Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models
Shuo Wang, Chihang Wang, Jia Gao +3
This study proposes a knowledge distillation algorithm based on large language models and feature alignment, aiming to effectively transfer the knowledge of large pre-trained model…
Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks
Jiacheng Hu, Xiaoxuan Liao, Jia Gao +3
This study proposes a large language model optimization method based on the improved LoRA fine-tuning algorithm, aiming to improve the accuracy and computational efficiency of the…
Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models
Zhen Qi, Jiajing Chen, Shuo Wang +3
This study aims to explore the performance improvement method of large language models based on GPT-4 under the multi-task learning framework and conducts experiments on two tasks:…
Few-Shot Learning with Adaptive Weight Masking in Conditional GANs
Jiacheng Hu, Zhen Qi, Jianjun Wei +3
Deep learning has revolutionized various fields, yet its efficacy is hindered by overfitting and the requirement of extensive annotated data, particularly in few-shot learning scen…
A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation
Jiajing Chen, Shuo Wang, Zhen Qi +3
This research introduces a novel text generation model that combines BERT's semantic interpretation strengths with GPT-4's generative capabilities, establishing a high standard in…
Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems
Jiajing Chen, Runyuan Bao, Hongye Zheng +3
This study aims to improve the accuracy and quality of large-scale language models (LLMs) in answering questions by integrating Elasticsearch into the Retrieval Augmented Generatio…