2 papers
cs.LG2025
DATA: Decomposed Attention-based Task Adaptation for Rehearsal-Free Continual Learning
Huanxuan Liao, Shizhu He, Yupu Hao +2
Continual learning (CL) is essential for Large Language Models (LLMs) to adapt to evolving real-world demands, yet they are susceptible to catastrophic forgetting (CF). While tradi…
cs.CL2024
CITI: Enhancing Tool Utilizing Ability in Large Language Models without Sacrificing General Performance
Yupu Hao, Pengfei Cao, Zhuoran Jin +4
Tool learning enables the Large Language Models (LLMs) to interact with the external environment by invoking tools, enriching the accuracy and capability scope of LLMs. However, pr…