Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
PromptSafe: Gated Prompt Tuning for Safe Text-to-Image Generation
Zonglei Jing, Xiao Yang, Xiaoqian Li +4
Text-to-image (T2I) models have demonstrated remarkable generative capabilities but remain vulnerable to producing not-safe-for-work (NSFW) content, such as violent or explicit ima…
cs.CV2025
How to Enable LLM with 3D Capacity? A Survey of Spatial Reasoning in LLM
Jirong Zha, Yuxuan Fan, Xiao Yang +2
3D spatial understanding is essential in real-world applications such as robotics, autonomous vehicles, virtual reality, and medical imaging. Recently, Large Language Models (LLMs)…