5 papers
Does Visual Information Play a Decisive Role in Vision-Language-Action Model Driving Behavior?
Jingtao He, Hongliang Lu, Xiaoyun Qiu +2
Vision-Language-Action (VLA) models have demonstrated promising capability in autonomous driving, highlighting the potential of unified multimodal architectures for jointly modelin…
Escaping the Linearity Trap: Manifold Detours for Black-Box Adversarial Attacks on Singing Audio Deepfake Detection
Yifan Liao, Yule Liu, Zhen Sun +5
Recent Singing Voice Synthesis (SVS) advances enable highly realistic but potentially malicious AI covers, making singing voice deepfake detection (SVDD) crucial. Self-Supervised L…
Stego Battlefield: Evaluating Image Steganography Attacks and Steganalysis Defenses
Zhen Sun, Zongmin Zhang, Leyi Sheng +7
Image steganography is widely used to protect user privacy and enable covert communication. However, it can also be abused by the adversary as a covert channel to bypass content mo…
Work Zones challenge VLM Trajectory Planning: Toward Mitigation and Robust Autonomous Driving
Yifan Liao, Zhen Sun, Xiaoyun Qiu +7
Visual Language Models (VLMs), with powerful multimodal reasoning capabilities, are gradually integrated into autonomous driving by several automobile manufacturers to enhance plan…
An Intention-driven Lane Change Framework Considering Heterogeneous Dynamic Cooperation in Mixed-traffic Environment
Xiaoyun Qiu, Haichao Liu, Yue Pan +2
In mixed-traffic environments, autonomous vehicles (AVs) must interact with heterogeneous human-driven vehicles (HVs) whose intentions and driving styles vary across individuals an…