collaborators

6 papers

cs.CL2024

Feature Alignment-Based Knowledge Distillation for Efficient Compression of Large Language Models

Shuo Wang, Chihang Wang, Jia Gao +3

This study proposes a knowledge distillation algorithm based on large language models and feature alignment, aiming to effectively transfer the knowledge of large pre-trained model…

cs.CL2024

Optimizing Large Language Models with an Enhanced LoRA Fine-Tuning Algorithm for Efficiency and Robustness in NLP Tasks

Jiacheng Hu, Xiaoxuan Liao, Jia Gao +3

This study proposes a large language model optimization method based on the improved LoRA fine-tuning algorithm, aiming to improve the accuracy and computational efficiency of the…

cs.CL2024

Optimizing Multi-Task Learning for Enhanced Performance in Large Language Models

Zhen Qi, Jiajing Chen, Shuo Wang +3

This study aims to explore the performance improvement method of large language models based on GPT-4 under the multi-task learning framework and conducts experiments on two tasks:…

cs.CV2024

Few-Shot Learning with Adaptive Weight Masking in Conditional GANs

Jiacheng Hu, Zhen Qi, Jianjun Wei +3

Deep learning has revolutionized various fields, yet its efficacy is hindered by overfitting and the requirement of extensive annotated data, particularly in few-shot learning scen…

cs.CL2024

A Combined Encoder and Transformer Approach for Coherent and High-Quality Text Generation

Jiajing Chen, Shuo Wang, Zhen Qi +3

This research introduces a novel text generation model that combines BERT's semantic interpretation strengths with GPT-4's generative capabilities, establishing a high standard in…

cs.IR2024

Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems

Jiajing Chen, Runyuan Bao, Hongye Zheng +3

This study aims to improve the accuracy and quality of large-scale language models (LLMs) in answering questions by integrating Elasticsearch into the Retrieval Augmented Generatio…