1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CL2026
LEAD: Layer-wise Expert-aligned Decoding for Faithful Radiology Report Generation
Ruixiao Yang, Yuanhe Tian, Xu Yang +2
Radiology Report Generation (RRG) aims to produce accurate and coherent diagnostics from medical images. Although large vision language models (LVLM) improve report fluency and acc…
cs.CV2025
MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs
Huiyi Chen, Jiawei Peng, Dehai Min +5
Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-world applications. However, exist…
cs.CV2024★ 1 cited
Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models
Yuhang Huang, Zihan Wu, Chongyang Gao +2
Large Vision-Language Models (LVLMs) are gaining traction for their remarkable ability to process and integrate visual and textual data. Despite their popularity, the capacity of L…