5 papers
Environmental Injection Attacks against GUI Agents in Realistic Dynamic Environments
Yitong Zhang, Ximo Li, Liyi Cai +1
Graphical User Interface (GUI) agents are increasingly deployed to interact with online web services, yet their exposure to open-world content renders them vulnerable to Environmen…
DAVSP: Safety Alignment for Large Vision-Language Models via Deep Aligned Visual Safety Prompt
Yitong Zhang, Jia Li, Liyi Cai +1
Large Vision-Language Models (LVLMs) have achieved impressive progress across various applications but remain vulnerable to malicious queries that exploit the visual modality. Exis…
Beyond Autoregression: An Empirical Study of Diffusion Large Language Models for Code Generation
Chengze Li, Yitong Zhang, Jia Li +2
LLMs have become the mainstream approaches to code generation. Existing LLMs mainly employ autoregressive generation, i.e. generating code token-by-token from left to right. Howeve…
AI-Driven Self-Evolving Software: A Promising Path Toward Software Automation
Liyi Cai, Yijie Ren, Yitong Zhang +1
Software automation has long been a central goal of software engineering, striving for software development that proceeds without human intervention. Recent efforts have leveraged…
A Few-Shot Metric Learning Method with Dual-Channel Attention for Cross-Modal Same-Neuron Identification
Wenwei Li, Liyi Cai, Wu Chen +1
In neuroscience research, achieving single-neuron matching across different imaging modalities is critical for understanding the relationship between neuronal structure and functio…