2 papers
cs.CR2024
Enhancing LLM-based Autonomous Driving Agents to Mitigate Perception Attacks
Ruoyu Song, Muslum Ozgur Ozmen, Hyungsub Kim +2
There is a growing interest in integrating Large Language Models (LLMs) with autonomous driving (AD) systems. However, AD systems are vulnerable to attacks against their object det…
cs.CL2024
Rethinking How to Evaluate Language Model Jailbreak
Hongyu Cai, Arjun Arunasalam, Leo Y. Lin +2
Large language models (LLMs) have become increasingly integrated with various applications. To ensure that LLMs do not generate unsafe responses, they are aligned with safeguards t…