From the 1 of 1 linked paper with an AI index.
1 citations · 1 across the 1 of their papers we have counts for
1 paper
Haozheng Luo, Jiahao Yu, Wenxin Zhang +9
The paper proposes a training-free, plug-and-play method that uses knowledge distillation and model fusion to correct misaligned (shadow-aligned) large language models, improving s…