activity
20132022
most citedMASS: Masked Sequence to Sequence Pre-training for Language Generation

578 citations · 2.7k across the 89 of their papers we have counts for

collaborators

136 papers

stat.ML2022

TD3 with Reverse KL Regularizer for Offline Reinforcement Learning from Mixed Datasets

Yuanying Cai, Chuheng Zhang, Li Zhao +6

We consider an offline reinforcement learning (RL) setting where the agent need to learn from a dataset collected by rolling out multiple behavior policies. There are two challenge…

q-bio.BM20222 cited

Incorporating Pre-training Paradigm for Antibody Sequence-Structure Co-design

Kaiyuan Gao, Lijun Wu, Jinhua Zhu +8

Antibodies are versatile proteins that can bind to pathogens and provide effective protection for human body. Recently, deep learning-based computational antibody design has attrac…

cs.SD202220 cited

Museformer: Transformer with Fine- and Coarse-Grained Attention for Music Generation

Botao Yu, Peiling Lu, Rui Wang +6

Symbolic music generation aims to generate music scores automatically. A recent trend is to use Transformer or its variants in music generation, which is, however, suboptimal, beca…

eess.AS202235 cited

NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality

Xu Tan, Jiawei Chen, Haohe Liu +11

Text to speech (TTS) has made rapid progress in both academia and industry in recent years. Some questions naturally arise that whether a TTS system can achieve human-level quality…

cs.LG202219 cited

METRO: Efficient Denoising Pretraining of Large Scale Autoencoding Language Models with Model Generated Signals

Payal Bajaj, Chenyan Xiong, Guolin Ke +7

We present an efficient method of pretraining large-scale autoencoding language models using training signals generated by an auxiliary model. Originated in ELECTRA, this training…

eess.AS2022

AdaSpeech 4: Adaptive Text to Speech in Zero-Shot Scenarios

Yihan Wu, Xu Tan, Bohan Li +5

Adaptive text to speech (TTS) can synthesize new voices in zero-shot scenarios efficiently, by using a well-trained source TTS model without adapting it on the speech data of new s…