Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Counterfactual Segmentation Reasoning: Diagnosing and Mitigating Pixel-Grounding Hallucination
Xinzhuo Li, Adheesh Juvekar, Jiaxun Zhang +6
Segmentation Vision-Language Models (VLMs) have significantly advanced grounded visual understanding, yet they remain prone to pixel-grounding hallucinations, producing masks for i…
cs.CV2025
INFELM: In-depth Fairness Evaluation of Large Text-To-Image Models
Di Jin, Xing Liu, Yu Liu +7
The rapid development of large language models (LLMs) and large vision models (LVMs) have propelled the evolution of multi-modal AI systems, which have demonstrated the remarkable…