2 papers
cs.CV2026
Can Multimodal Large Language Models Understand OCT?
Baochen Fu, Wenzhi Deng, Baihao Jin +5
Optical coherence tomography (OCT) imaging is essential for the diagnosis and treatment of retinal diseases. Although multimodal large language models (MLLMs) have demonstrated con…
cs.CL2026
MMKU-Bench: A Multimodal Update Benchmark for Diverse Visual Knowledge
Baochen Fu, Yuntao Du, Cheng Chang +6
As real-world knowledge continues to evolve, the parametric knowledge acquired by multimodal models during pretraining becomes increasingly difficult to remain consistent with real…