1 paper
Yuquan Wang, Mi Zhang, Yining Wang +4
Large Reasoning Models (LRMs) have demonstrated impressive performance in reasoning-intensive tasks, but they remain vulnerable to harmful content generation, particularly in the m…