3 papers
cs.CL2025
FALCON: Fine-grained Activation Manipulation by Contrastive Orthogonal Unalignment for Large Language Model
Jinwei Hu, Zhenglin Huang, Xiangyu Yin +4
Large language models have been widely applied, but can inadvertently encode sensitive or harmful information, raising significant safety concerns. Machine unlearning has emerged t…
cs.LG2025
Trustworthy Text-to-Image Diffusion Models: A Timely and Focused Survey
Yi Zhang, Zhen Chen, Chih-Hong Cheng +6
Text-to-Image (T2I) Diffusion Models (DMs) have garnered widespread attention for their impressive advancements in image generation. However, their growing popularity has raised et…
cs.CV2025
A Black-Box Evaluation Framework for Semantic Robustness in Bird's Eye View Detection
Fu Wang, Yanghao Zhang, Xiangyu Yin +4
Camera-based Bird's Eye View (BEV) perception models receive increasing attention for their crucial role in autonomous driving, a domain where concerns about the robustness and rel…