3 papers
cs.AI2026
CLEAR: Continuous Latent Adapter Routing for Utility-Preserving LLM Safety Alignment
Chengxiao Wang, Enyi Jiang, Xiaojing Liao +1
Improving the safety of large language models (LLMs) often comes at the expense of utility, as globally applied safety tuning may affect model responses to both harmful and benign…
cs.AI2025
Robust Answers, Fragile Logic: Probing the Decoupling Hypothesis in LLM Reasoning
Enyi Jiang, Changming Xu, Nischay Singh +2
While Chain-of-Thought (CoT) prompting has become a cornerstone for complex reasoning in Large Language Models (LLMs), the faithfulness of the generated reasoning remains an open q…
cs.LG2024
Towards Generalized Certified Robustness with Multi-Norm Training
Enyi Jiang, David S. Cheung, Gagandeep Singh
Existing certified training methods can only train models to be robust against a certain perturbation type (e.g. or ). However, an certifiably robust mod…