7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 7 cited
CogVLM2: Visual Language Models for Image and Video Understanding
Wenyi Hong, Weihan Wang, Ming Ding +22
Beginning with VisualGLM and CogVLM, we are continuously exploring VLMs in pursuit of enhanced vision-language fusion, efficient higher-resolution architecture, and broader modalit…
cs.DC2023
Pipeline MoE: A Flexible MoE Implementation with Pipeline Parallelism
Xin Chen, Hengheng Zhang, Xiaotao Gu +3
The Mixture of Experts (MoE) model becomes an important choice of large language models nowadays because of its scalability with sublinear computational complexity for training and…