2 papers
cs.CV2026
Two Sides of the Same Coin: Co-Evolving Search for Cross-Task Attacks on Vision-Language Models
Xuanhui Lin, Junhao Dong, Mingrong Gong +3
Vision-language models (VLMs) exhibit strong generalization across multimodal tasks but remain vulnerable to adversarial perturbations. Existing attacks typically follow single-tra…
cs.SD2026
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook
Kaiwen Luo, Zhenhong Zhou, Leo Wang +34
Advances in Large Language Models (LLMs) have paved the way for Multimodal Large Language Models (MLLMs). Among these, Large Audio Language Models (LALMs) are essential for realizi…