3 papers
cs.LG2025
Does Low Rank Adaptation Lead to Lower Robustness against Training-Time Attacks?
Zi Liang, Haibo Hu, Qingqing Ye +2
Low rank adaptation (LoRA) has emerged as a prominent technique for fine-tuning large language models (LLMs) thanks to its superb efficiency gains over previous methods. While exte…
cs.CL2023
Healing Unsafe Dialogue Responses with Weak Supervision Signals
Zi Liang, Pinghui Wang, Ruofei Zhang +3
Recent years have seen increasing concerns about the unsafe response generation of large-scale dialogue systems, where agents will learn offensive or biased behaviors from the real…
cs.CL2023
Multi-Action Dialog Policy Learning from Logged User Feedback
Shuo Zhang, Junzhou Zhao, Pinghui Wang +5
Multi-action dialog policy, which generates multiple atomic dialog actions per turn, has been widely applied in task-oriented dialog systems to provide expressive and efficient sys…