activity
20172022
most citedInvolution: Inverting the Inherence of Convolution for Visual Recognition

25 citations · 70 across the 16 of their papers we have counts for

collaborators

43 papers

cs.CL20221 cited

ExtremeBERT: A Toolkit for Accelerating Pretraining of Customized BERT

Rui Pan, Shizhe Diao, Jianlin Chen +1

In this paper, we present ExtremeBERT, a toolkit for accelerating and customizing BERT pretraining. Our goal is to provide an easy-to-use BERT pretraining toolkit for the research…

cs.LG2022

Normalizing Flow with Variational Latent Representation

Hanze Dong, Shizhe Diao, Weizhong Zhang +1

Normalizing flow (NF) has gained popularity over traditional maximum likelihood based methods due to its strong capability to model complex data distributions. However, the standar…

cs.CV20221 cited

FAF: A novel multimodal emotion recognition approach integrating face, body and text

Zhongyu Fang, Aoyun He, Qihui Yu +4

Multimodal emotion analysis performed better in emotion recognition depending on more comprehensive emotional clues and multimodal emotion dataset. In this paper, we developed a la…

physics.acc-ph2022

Prior-mean-assisted Bayesian optimization application on FRIB Front-End tunning

Kilean Hwang, Tomofumi Maruta, Alexander Plastun +5

Bayesian optimization~(BO) is often used for accelerator tuning due to its high sample efficiency. However, the computational scalability of training over large data-set can be pro…

cs.CL20222 cited

MICO: A Multi-alternative Contrastive Learning Framework for Commonsense Knowledge Representation

Ying Su, Zihao Wang, Tianqing Fang +3

Commonsense reasoning tasks such as commonsense knowledge graph completion and commonsense question answering require powerful representation learning. In this paper, we propose to…

cs.LG2022

A Self-Play Posterior Sampling Algorithm for Zero-Sum Markov Games

Wei Xiong, Han Zhong, Chengshuai Shi +2

Existing studies on provably efficient algorithms for Markov games (MGs) almost exclusively build on the "optimism in the face of uncertainty" (OFU) principle. This work focuses on…