3 papers
cs.LG2026
Pay Less Attention to Function Words for Free Robustness of Vision-Language Models
Qiwei Tian, Chenhao Lin, Zhengyu Zhao +1
To address the trade-off between robustness and performance for robust VLM, we observe that function words could incur vulnerability of VLMs against cross-modal adversarial attacks…
cs.CV2025
HCMA: Hierarchical Cross-model Alignment for Grounded Text-to-Image Generation
Hang Wang, Zhi-Qi Cheng, Chenhao Lin +2
Text-to-image synthesis has progressed to the point where models can generate visually compelling images from natural language prompts. Yet, existing methods often fail to reconcil…
cs.SE2025
Deep Learning Library Testing: Definition, Methods and Challenges
Xiaoyu Zhang, Weipeng Jiang, Chao Shen +4
In recent years, software systems powered by deep learning (DL) techniques have significantly facilitated people's lives in many aspects. As the backbone of these DL systems, vario…