97 citations · 115 across the 10 of their papers we have counts for
Showing 2023 · cs.CVShow all
3 papers · 2 filters
cs.CV2023★ 1 cited
What Matters in Training a GPT4-Style Language Model with Multimodal Inputs?
Yan Zeng, Hanbo Zhang, Jiani Zheng +5
Recent advancements in Large Language Models (LLMs) such as GPT4 have displayed exceptional multi-modal capabilities in following open-ended instructions given images. However, the…
cs.CV2023★ 1 cited
eTag: Class-Incremental Learning with Embedding Distillation and Task-Oriented Generation
Libo Huang, Yan Zeng, Chuanguang Yang +3
Class-Incremental Learning (CIL) aims to solve the neural networks' catastrophic forgetting problem, which refers to the fact that once the network updates on a new task, its perfo…
cs.CV2023
Toward Building General Foundation Models for Language, Vision, and Vision-Language Understanding Tasks
Xinsong Zhang, Yan Zeng, Jipeng Zhang +1
Foundation models or pre-trained models have substantially improved the performance of various language, vision, and vision-language understanding tasks. However, existing foundati…