3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CL2024
A Learning Rate Path Switching Training Paradigm for Version Updates of Large Language Models
Zhihao Wang, Shiyu Liu, Jianheng Huang +5
Due to the continuous emergence of new data, version updates have become an indispensable requirement for Large Language Models (LLMs). The training paradigms for version updates o…
cs.CL2024★ 3 cited
Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
Jianheng Huang, Leyang Cui, Ante Wang +5
Large language models (LLMs) suffer from catastrophic forgetting during continual learning. Conventional rehearsal-based methods rely on previous training data to retain the model'…
cs.CL2023
Response Enhanced Semi-supervised Dialogue Query Generation
Jianheng Huang, Ante Wang, Linfeng Gao +2
Leveraging vast and continually updated knowledge from the Internet has been considered an important ability for a dialogue system. Therefore, the dialogue query generation task is…