Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
Sifan Li, Hongkai Chen, Yujun Cai +4
Vision Language Models (VLMs) have achieved impressive progress in multimodal reasoning; yet, they remain vulnerable to hallucinations, where outputs are not grounded in visual evi…
cs.CV2025
Lost in Edits? A -Compass for AIGC Provenance
Wenhao You, Bryan Hooi, Yiwei Wang +5
Recent advancements in diffusion models have driven the growth of text-guided image editing tools, enabling precise and iterative modifications of synthesized content. However, as…