2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 2 cited
CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts
Jiachen Li, Xinyao Wang, Sijie Zhu +6
Recent advancements in Multimodal Large Language Models (LLMs) have focused primarily on scaling by increasing text-image pair data and enhancing LLMs to improve performance on mul…
cs.CV2023
Video Instance Matting
Jiachen Li, Roberto Henschel, Vidit Goel +3
Conventional video matting outputs one alpha matte for all instances appearing in a video frame so that individual instances are not distinguished. While video instance segmentatio…