3 citations · 8 across the 19 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.CV2024★ 3 cited
Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs
Xiaofeng Zhang, Yihao Quan, Chaochen Gu +6
The hallucination problem in multimodal large language models (MLLMs) remains a common issue. Although image tokens occupy a majority of the input sequence of MLLMs, there is limit…
cs.CL2024★ 3 cited
Instance-adaptive Zero-shot Chain-of-Thought Prompting
Xiaosong Yuan, Chen Shen, Shaotian Yan +6
Zero-shot Chain-of-Thought (CoT) prompting emerges as a simple and effective strategy for enhancing the performance of large language models (LLMs) in real-world reasoning tasks. N…
cs.CL2024★ 2 cited
From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks
Xiaofeng Zhang, Yihao Quan, Chen Shen +7
Large Vision Language Models (LVLMs) achieve great performance on visual-language reasoning tasks, however, the black-box nature of LVLMs hinders in-depth research on the reasoning…