Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention
Wenbin An, Feng Tian, Sicong Leng +6
Despite great success across various multimodal tasks, Large Vision-Language Models (LVLMs) often encounter object hallucinations with generated textual responses being inconsisten…
cs.CV2024
A Review of Multimodal Explainable Artificial Intelligence: Past, Present and Future
Shilin Sun, Wenbin An, Feng Tian +5
Artificial intelligence (AI) has rapidly developed through advancements in computational power and the growth of massive datasets. However, this progress has also heightened challe…