Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Learning to Edit Knowledge via Instruction-based Chain-of-Thought Prompting
Jinhu Fu, Yan Bai, Longzhu He +4
Large language models (LLMs) can effectively handle outdated information through knowledge editing. However, current approaches face two key limitations: (I) Poor generalization: M…
cs.CL2024
Alignment-Enhanced Decoding:Defending via Token-Level Adaptive Refining of Probability Distributions
Quan Liu, Zhenhong Zhou, Longzhu He +3
Large language models are susceptible to jailbreak attacks, which can result in the generation of harmful content. While prior defenses mitigate these risks by perturbing or inspec…