4 papers
MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution
Zefeng Wang, Minxi Yan, Jinhe Bi +3
Recent LLM agents tackle increasingly long-horizon, open-ended tasks, and external skills, reusable procedural knowledge supplied to the agent, further extend this capability. Howe…
ICM-Fusion: In-Context Meta-Optimized LoRA Fusion for Multi-Task Adaptation
Yihua Shao, Xiaofeng Lin, Xinwei Long +7
Enabling multi-task adaptation in pre-trained Low-Rank Adaptation (LoRA) models is crucial for enhancing their generalization capabilities. Most existing pre-trained LoRA fusion me…
In-Context Meta LoRA Generation
Yihua Shao, Minxi Yan, Yang Liu +12
Low-rank Adaptation (LoRA) has demonstrated remarkable capabilities for task specific fine-tuning. However, in scenarios that involve multiple tasks, training a separate LoRA model…
GWQ: Gradient-Aware Weight Quantization for Large Language Models
Yihua Shao, Yan Gu, Siyu Chen +12
Large language models (LLMs) show impressive performance in solving complex language tasks. However, its large number of parameters presents significant challenges for the deployme…