2 citations · 5 across the 9 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.CV2024★ 1 cited
EE-MLLM: A Data-Efficient and Compute-Efficient Multimodal Large Language Model
Feipeng Ma, Yizhou Zhou, Zheyu Zhang +7
Recent advancements in Multimodal Large Language Models (MLLMs) have demonstrated satisfactory performance across various vision-language tasks. Current approaches for vision and l…
cs.CV2024★ 1 cited
Visual Perception by Large Language Model's Weights
Feipeng Ma, Hongwei Xue, Guangting Wang +7
Existing Multimodal Large Language Models (MLLMs) follow the paradigm that perceives visual information by aligning visual features with the input space of Large Language Models (L…
cs.CV2024
Multi-Modal Generative Embedding Model
Feipeng Ma, Hongwei Xue, Guangting Wang +7
Most multi-modal tasks can be formulated into problems of either generation or embedding. Existing models usually tackle these two types of problems by decoupling language modules…