activity
20192022
most citedLabel-Wise Document Pre-Training for Multi-Label Text Classification

2 citations · 9 across the 11 of their papers we have counts for

collaborators

13 papers

cs.CV2022

Towards Unifying Reference Expression Generation and Comprehension

Duo Zheng, Tao Kong, Ya Jing +2

Reference Expression Generation (REG) and Comprehension (REC) are two highly correlated tasks. Modeling REG and REC simultaneously for utilizing the relation between them is a prom…

cs.CV20221 cited

Question-Driven Graph Fusion Network For Visual Question Answering

Yuxi Qian, Yuncong Hu, Ruonan Wang +2

Existing Visual Question Answering (VQA) models have explored various visual relationships between objects in the image to answer complex questions, which inevitably introduces irr…

cs.CL20221 cited

Co-VQA : Answering by Interactive Sub Question Sequence

Ruonan Wang, Yuxi Qian, Fangxiang Feng +2

Most existing approaches to Visual Question Answering (VQA) answer questions directly, however, people usually decompose a complex question into a sequence of simple sub questions…

cs.CV20221 cited

Spot the Difference: A Cooperative Object-Referring Game in Non-Perfectly Co-Observable Scene

Duo Zheng, Fandong Meng, Qingyi Si +5

Visual dialog has witnessed great progress after introducing various vision-oriented goals into the conversation, especially such as GuessWhich and GuessWhat, where the only image…

cs.IR2021

Deep Keyphrase Completion

Yu Zhao, Jia Song, Huali Feng +4

Keyphrase provides accurate information of document content that is highly compact, concise, full of meanings, and widely used for discourse comprehension, organization, and text r…

cs.CL2021

Converse, Focus and Guess -- Towards Multi-Document Driven Dialogue

Han Liu, Caixia Yuan, Xiaojie Wang +3

We propose a novel task, Multi-Document Driven Dialogue (MD3), in which an agent can guess the target document that the user is interested in by leading a dialogue. To benchmark pr…