activity
20192022
most citedContrastive Visual-Linguistic Pretraining

18 citations · 22 across the 6 of their papers we have counts for

collaborators

7 papers

cs.CL2022

Beyond the Granularity: Multi-Perspective Dialogue Collaborative Selection for Dialogue State Tracking

Jinyu Guo, Kai Shuang, Jijie Li +2

In dialogue state tracking, dialogue history is a crucial material, and its utilization varies between different models. However, no matter how the dialogue history is used, each e…

cs.CV2021

Dense Contrastive Visual-Linguistic Pretraining

Lei Shi, Kai Shuang, Shijie Geng +5

Inspired by the success of BERT, several multimodal representation learning approaches have been proposed that jointly represent image and text. These approaches achieve superior p…

cs.CL2021

Dual Slot Selector via Local Reliability Verification for Dialogue State Tracking

Jinyu Guo, Kai Shuang, Jijie Li +1

The goal of dialogue state tracking (DST) is to predict the current dialogue state given all previous dialogue contexts. Existing approaches generally predict the dialogue state at…

cs.LG20201 cited

A Hierarchical User Intention-Habit Extract Network for Credit Loan Overdue Risk Detection

Hao Guo, Xintao Ren, Rongrong Wang +3

More personal consumer loan products are emerging in mobile banking APP. For ease of use, application process is always simple, which means that few application information is requ…

cs.CV202018 cited

Contrastive Visual-Linguistic Pretraining

Lei Shi, Kai Shuang, Shijie Geng +6

Several multi-modality representation learning approaches such as LXMERT and ViLBERT have been proposed recently. Such approaches can achieve superior performance due to the high-l…

cs.CV20203 cited

Multi-Layer Content Interaction Through Quaternion Product For Visual Question Answering

Lei Shi, Shijie Geng, Kai Shuang +4

Multi-modality fusion technologies have greatly improved the performance of neural network-based Video Description/Caption, Visual Question Answering (VQA) and Audio Visual Scene-a…