1 citations · 1 across the 2 of their papers we have counts for
4 papers
Taming a Retrieval Framework to Read Images in Humanlike Manner for Augmenting Generation of MLLMs
Suyang Xi, Chenxi Yang, Hong Ding +4
Multimodal large language models (MLLMs) often fail in fine-grained visual question answering, producing hallucinations about object identities, positions, and relations because te…
Copula-enhanced Vision Transformer for high myopia diagnosis through OU UWF fundus images
Chong Zhong, Yunhao Liu, Yang Li +10
The advancement of AI-assisted myopia screening necessitates the joint diagnosis of both-eye (OU) high myopia (HM) status and the prediction of axial length (AL). This clinical req…
On MCMC mixing for predictive inference under unidentified transformation models
Chong Zhong, Jin Yang, Junshan Shen +2
Reliable Bayesian predictive inference has long been an open problem under unidentified transformation models, since the Markov Chain Monte Carlo (MCMC) chains of posterior predict…
OU-CoViT: Copula-Enhanced Bi-Channel Multi-Task Vision Transformers with Dual Adaptation for OU-UWF Images
Yang Li, Jianing Deng, Chong Zhong +7
Myopia screening using cutting-edge ultra-widefield (UWF) fundus imaging and joint modeling of multiple discrete and continuous clinical scores presents a promising new paradigm fo…