2 papers
cs.AI2025
Value-Aligned Prompt Moderation via Zero-Shot Agentic Rewriting for Safe Image Generation
Xin Zhao, Xiaojun Chen, Bingshan Liu +3
Generative vision-language models like Stable Diffusion demonstrate remarkable capabilities in creative media synthesis, but they also pose substantial risks of producing unsafe, o…
cs.CR2025
Who Speaks for the Trigger? Dynamic Expert Routing in Backdoored Mixture-of-Experts Transformers
Xin Zhao, Xiaojun Chen, Bingshan Liu +3
Large language models (LLMs) with Mixture-of-Experts (MoE) architectures achieve impressive performance and efficiency by dynamically routing inputs to specialized subnetworks, kno…