9 citations · 9 across the 2 of their papers we have counts for
3 papers
cs.AI2025
Self-Aware Safety Augmentation: Leveraging Internal Semantic Understanding to Enhance Safety in Vision-Language Models
Wanying Wang, Zeyu Ma, Han Zheng +2
Large vision-language models (LVLMs) are vulnerable to harmful input compared to their language-only backbones. We investigated this vulnerability by exploring LVLMs internal dynam…
cs.CV2024★ 9 cited
PIG: Prompt Images Guidance for Night-Time Scene Parsing
Zhifeng Xie, Rui Qiu, Sen Wang +3
Night-time scene parsing aims to extract pixel-level semantic information in night images, aiding downstream tasks in understanding scene object distribution. Due to limited labele…
cs.MM2024
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
Sen Wang, Jiangning Zhang, Xin Tan +3
The body movements accompanying speech aid speakers in expressing their ideas. Co-speech motion generation is one of the important approaches for synthesizing realistic avatars. Du…