1 paper
Ke Miao, Jiaxin Li, Hongliang Chen +2
While Large Reasoning Models (LRMs) excel at complex tasks, they remain highly vulnerable to sophisticated jailbreaks and direct harmful queries. To address this vulnerability, pri…