37 citations · 41 across the 3 of their papers we have counts for
3 papers
cs.CV2021
A Picture is Worth a Thousand Words: A Unified System for Diverse Captions and Rich Images Generation
Yupan Huang, Bei Liu, Jianlong Fu +1
A creative image-and-text generative AI system mimics humans' extraordinary abilities to provide users with diverse and comprehensive caption suggestions, as well as rich image cre…
cs.CV2021★ 37 cited
Unifying Multimodal Transformer for Bi-directional Image and Text Generation
Yupan Huang, Hongwei Xue, Bei Liu +1
We study the joint learning of image-to-text and text-to-image generations, which are naturally bi-directional tasks. Typical existing works design two separate task-specific model…
cs.CV2019★ 4 cited
Decoupling Localization and Classification in Single Shot Temporal Action Detection
Yupan Huang, Qi Dai, Yutong Lu
Video temporal action detection aims to temporally localize and recognize the action in untrimmed videos. Existing one-stage approaches mostly focus on unifying two subtasks, i.e.,…