65 citations · 65 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 65 cited
MultiModal-GPT: A Vision and Language Model for Dialogue with Humans
Tao Gong, Chengqi Lyu, Shilong Zhang +7
We present a vision and language model named MultiModal-GPT to conduct multi-round dialogue with humans. MultiModal-GPT can follow various instructions from humans, such as generat…
cs.CV2021
Multi-Semantic Image Recognition Model and Evaluating Index for explaining the deep learning models
Qianmengke Zhao, Ye Wang, Qun Liu
Although deep learning models are powerful among various applications, most deep learning models are still a black box, lacking verifiability and interpretability, which means the…