2 papers
cs.CV2025
OViP: Online Vision-Language Preference Learning for VLM Hallucination
Shujun Liu, Siyuan Wang, Zejun Li +3
Large vision-language models (LVLMs) remain vulnerable to hallucination, often generating content misaligned with visual inputs. Although recent training-based approaches aim to mi…
cs.CV2025
NSFW-Classifier Guided Prompt Sanitization for Safe Text-to-Image Generation
Yu Xie, Chengjie Zeng, Lingyun Zhang +1
The rapid advancement of text-to-image (T2I) models, such as Stable Diffusion, has enhanced their capability to synthesize images from textual prompts. However, this progress also…