17 citations · 53 across the 26 of their papers we have counts for
11 papers · 1 filter
Controlling Character Motions without Observable Driving Source
Weiyuan Li, Bin Dai, Ziyi Zhou +2
How to generate diverse, life-like, and unlimited long head/body sequences without any driving source? We argue that this under-investigated research problem is non-trivial at all,…
Reinforced Disentanglement for Face Swapping without Skip Connection
Xiaohang Ren, Xingyu Chen, Pengfei Yao +2
The SOTA face swap models still suffer the problem of either target identity (i.e., shape) being leaked or the target non-identity attributes (i.e., background, hair) failing to be…
Visual Instruction Tuning with Polite Flamingo
Delong Chen, Jianfeng Liu, Wenliang Dai +1
Recent research has demonstrated that the multi-task fine-tuning of multi-modal Large Language Models (LLMs) using an assortment of annotated downstream vision-language datasets si…
LiveChat: A Large-Scale Personalized Dialogue Dataset Automatically Constructed from Live Streaming
Jingsheng Gao, Yixin Lian, Ziyi Zhou +2
Open-domain dialogue systems have made promising progress in recent years. While the state-of-the-art dialogue agents are built upon large-scale text-based social media data and la…
ViTMatte: Boosting Image Matting with Pretrained Plain Vision Transformers
Jingfeng Yao, Xinggang Wang, Shusheng Yang +1
Recently, plain vision Transformers (ViTs) have shown impressive performance on various computer vision tasks, thanks to their strong modeling capacity and large-scale pretraining.…
Hierarchical Verbalizer for Few-Shot Hierarchical Text Classification
Ke Ji, Yixin Lian, Jingsheng Gao +1
Due to the complex label hierarchy and intensive labeling cost in practice, the hierarchical text classification (HTC) suffers a poor performance especially when low-resource or fe…