Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
MoEGen: Mixture-of-Experts for Instance-Adaptive LoRA Generation
Yiming Zeng, Lei Lu, Zexin Li +9
Parameter-efficient fine-tuning (PEFT) enables efficient adaptation of large language models, but existing MoE-based PEFT methods typically improve capacity by storing multiple ful…
cs.CL2024
All-in-One Tuning and Structural Pruning for Domain-Specific LLMs
Lei Lu, Zhepeng Wang, Runxue Bao +7
Existing pruning techniques for large language models (LLMs) targeting domain-specific applications typically follow a two-stage process: pruning the pretrained general-purpose LLM…