most citedCoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP

11 citations · 21 across the 5 of their papers we have counts for

collaborators

6 papers

cs.CV20225 cited

ImaginaryNet: Learning Object Detectors without Real Images and Annotations

Minheng Ni, Zitong Huang, Kailai Feng +1

Without the demand of training in reality, humans can easily detect a known concept simply based on its language description. Empowering deep learning with this ability undoubtedly…

cs.CV20221 cited

NÜWA-LIP: Language Guided Image Inpainting with Defect-free VQGAN

Minheng Ni, Chenfei Wu, Haoyang Huang +3

Language guided image inpainting aims to fill in the defective regions of an image under the guidance of text while keeping non-defective regions unchanged. However, the encoding p…

cs.CL20203 cited

Co-GAT: A Co-Interactive Graph Attention Network for Joint Dialog Act Recognition and Sentiment Classification

Libo Qin, Zhouyang Li, Wanxiang Che +2

In a dialog system, dialog act recognition and sentiment classification are two correlative tasks to capture speakers intentions, where dialog act and sentiment can indicate the ex…

cs.CL20201 cited

DCR-Net: A Deep Co-Interactive Relation Network for Joint Dialog Act Recognition and Sentiment Classification

Libo Qin, Wanxiang Che, Yangming Li +2

In dialog system, dialog act recognition and sentiment classification are two correlative tasks to capture speakers intentions, where dialog act and sentiment can indicate the expl…

cs.CL202011 cited

CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual NLP

Libo Qin, Minheng Ni, Yue Zhang +1

Multi-lingual contextualized embeddings, such as multilingual-BERT (mBERT), have shown success in a variety of zero-shot cross-lingual tasks. However, these models are limited by h…

cs.CL2020

M3P: Learning Universal Representations via Multitask Multilingual Multimodal Pre-training

Minheng Ni, Haoyang Huang, Lin Su +6

We present M3P, a Multitask Multilingual Multimodal Pre-trained model that combines multilingual pre-training and multimodal pre-training into a unified framework via multitask pre…