2 papers
cs.CV2026
ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs
Zitong Xu, Huiyu Duan, Shengyao Qin +6
Recent advances in multimodal large language models (MLLMs) have greatly improved image understanding and captioning capabilities. However, existing image captioning benchmarks typ…
cs.CV2025
VRS-UIE: Value-Driven Reordering Scanning for Underwater Image Enhancement
Kui Jiang, Yan Luo, Junjun Jiang +3
State Space Models (SSMs) have emerged as a promising backbone for vision tasks due to their linear complexity and global receptive field. However, in the context of Underwater Ima…