3 papers
cs.SE2026
A Comprehensive Study of Implementation Bugs in Multi-modal Agents
Suwan Li, Lei Bu, Shangqing Liu +5
Multi-Modal Agents (M-agents), empowered by Large Language Models (LLMs), excel in various complex, open-world scenarios such as autonomous driving and robotics. However, their uni…
cs.CV2025
SAP-DIFF: Semantic Adversarial Patch Generation for Black-Box Face Recognition Models via Diffusion Models
Mingsi Wang, Shuaiyin Yao, Chang Yue +2
Given the need to evaluate the robustness of face recognition (FR) models, many efforts have focused on adversarial patch attacks that mislead FR models by introducing localized pe…
cs.SE2024
Model-Enhanced LLM-Driven VUI Testing of VPA Apps
Suwan Li, Lei Bu, Guangdong Bai +3
The flourishing ecosystem centered around voice personal assistants (VPA), such as Amazon Alexa, has led to the booming of VPA apps. The largest app market Amazon skills store, for…