2 papers
cs.AI2026
Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers
Zihao Xue, Yan Wang, Zhen Bi +7
Diffusion Transformers have become a powerful backbone for text-to-image generation, but their layered and cross-modal generation process makes safety control fundamentally differe…
cs.LG2026
Thought Purity: A Defense Framework For Chain-of-Thought Attack
Zihao Xue, Zhen Bi, Long Ma +7
Large Reasoning Models (LRMs) leverage Chain-of-Thought (CoT) reasoning to solve complex tasks, but this explicit reasoning process introduces a critical vulnerability: adversarial…