3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2024
Dialectical Alignment: Resolving the Tension of 3H and Security Threats of LLMs
Shu Yang, Jiayuan Su, Han Jiang +5
With the rise of large language models (LLMs), ensuring they embody the principles of being helpful, honest, and harmless (3H), known as Human Alignment, becomes crucial. While exi…
cs.CL2024★ 3 cited
MoRAL: MoE Augmented LoRA for LLMs' Lifelong Learning
Shu Yang, Muhammad Asif Ali, Cheng-Long Wang +2
Adapting large language models (LLMs) to new domains/tasks and enabling them to be efficient lifelong learners is a pivotal challenge. In this paper, we propose MoRAL, i.e., Mixtur…