1 paper
Jiasen Zheng, Huajun Zhang, Xu Yan +2
This paper addresses the limitations of large-scale language models in safety alignment and robustness by proposing a fine-tuning method that combines contrastive distillation with…