3 papers
cs.LG2025
Feature-Selective Representation Misdirection for Machine Unlearning
Taozhao Chen, Linghan Huang, Kim-Kwang Raymond Choo +1
As large language models (LLMs) are increasingly adopted in safety-critical and regulated sectors, the retention of sensitive or prohibited knowledge introduces escalating risks, r…
cs.RO2025
Trust in LLM-controlled Robotics: a Survey of Security Threats, Defenses and Challenges
Xinyu Huang, Shyam Karthick V B, Taozhao Chen +5
The integration of Large Language Models (LLMs) into robotics has revolutionized their ability to interpret complex human commands and execute sophisticated tasks. However, such pa…
cs.CL2025
The Tower of Babel Revisited: Multilingual Jailbreak Prompts on Closed-Source Large Language Models
Linghan Huang, Haolin Jin, Zhaoge Bi +6
Large language models (LLMs) have seen widespread applications across various domains, yet remain vulnerable to adversarial prompt injections. While most existing research on jailb…