2 citations · 2 across the 2 of their papers we have counts for
4 papers
TPC: Cross-Temporal Prediction Connection for Vision-Language Model Hallucination Reduction
Chao Wang, Weiwei Fu, Yang Zhou
Vision-language models (VLMs) have achieved remarkable advancements, capitalizing on the impressive capabilities of large language models (LLMs) across diverse tasks. Despite this,…
Mitigating Hallucinations in Large Vision-Language Models with Internal Fact-based Contrastive Decoding
Chao Wang, Xuancheng Zhou, Weiwei Fu +1
Large Visual Language Models (LVLMs) integrate visual and linguistic modalities, exhibiting exceptional performance across various multimodal tasks. Nevertheless, LVLMs remain vuln…
MINT: Mitigating Hallucinations in Large Vision-Language Models via Token Reduction
Chao Wang, Jianming Yang, Yang Zhou
Hallucination has been a long-standing and inevitable problem that hinders the application of Large Vision-Language Models (LVLMs) in domains that require high reliability. Various…
VIKSER: Visual Knowledge-Driven Self-Reinforcing Reasoning Framework
Chao Wang, Chunbai Zhang, Yongxiao Tian +2
Visual reasoning refers to the task of solving questions about visual information. Current visual reasoning methods typically employ pre-trained vision-language model (VLM) strateg…