178 citations · 343 across the 10 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2022★ 3 cited
OFASys: A Multi-Modal Multi-Task Learning System for Building Generalist Models
Jinze Bai, Rui Men, Hao Yang +15
Generalist models, which are capable of performing diverse multi-modal tasks in a task-agnostic way within a single model, have been explored recently. Being, hopefully, an alterna…
cs.CV2021★ 3 cited
Connecting Language and Vision for Natural Language-Based Vehicle Retrieval
Shuai Bai, Zhedong Zheng, Xiaohan Wang +5
Vehicle search is one basic task for the efficient traffic management in terms of the AI City. Most existing practices focus on the image-based vehicle matching, including vehicle…
cs.CV2021
CogView: Mastering Text-to-Image Generation via Transformers
Ming Ding, Zhuoyi Yang, Wenyi Hong +8
Text-to-Image generation in the general domain has long been an open problem, which requires both a powerful generative model and cross-modal understanding. We propose CogView, a 4…