2 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
Kimi-VL Technical Report
Kimi Team, Angang Du, Bohong Yin +92
We present Kimi-VL, an efficient open-source Mixture-of-Experts (MoE) vision-language model (VLM) that offers advanced multimodal reasoning, long-context understanding, and strong…
cs.AI2023★ 2 cited
What Makes for Robust Multi-Modal Models in the Face of Missing Modalities?
Siting Li, Chenzhuang Du, Yue Zhao +2
With the growing success of multi-modal learning, research on the robustness of multi-modal models, especially when facing situations with missing modalities, is receiving increase…
cs.CV2023
Improving Discriminative Multi-Modal Learning with Large-Scale Pre-Trained Models
Chenzhuang Du, Yue Zhao, Chonghua Liao +3
This paper investigates how to better leverage large-scale pre-trained uni-modal models to further enhance discriminative multi-modal learning. Even when fine-tuned with only uni-m…