activity
20152022
most citedMultilingual Denoising Pre-training for Neural Machine Translation

607 citations · 1.5k across the 22 of their papers we have counts for

collaborators

43 papers

cs.CV20222 cited

f-DM: A Multi-stage Diffusion Model via Progressive Signal Transformation

Jiatao Gu, Shuangfei Zhai, Yizhe Zhang +2

Diffusion models (DMs) have recently emerged as SoTA tools for generative modeling in various domains. Standard DMs can be viewed as an instantiation of hierarchical variational au…

cs.DC20223 cited

Migrating from Microservices to Serverless: An IoT Platform Case Study

Mohak Chadha, Victor Pacyna, Anshul Jindal +2

Microservice architecture is the common choice for developing cloud applications these days since each individual microservice can be independently modified, replaced, and scaled.…

cs.SE20222 cited

Muffin: Testing Deep Learning Libraries via Neural Architecture Fuzzing

Jiazhen Gu, Xuchuan Luo, Yangfan Zhou +1

Deep learning (DL) techniques are proven effective in many challenging tasks, and become widely-adopted in practice. However, previous work has shown that DL libraries, the basis o…

cs.CL2022

Detection, Disambiguation, Re-ranking: Autoregressive Entity Linking as a Multi-Task Problem

Khalil Mrini, Shaoliang Nie, Jiatao Gu +3

We propose an autoregressive entity linking model, that is trained with two auxiliary tasks, and learns to re-rank generated samples at inference time. Our proposed novelties addre…

cs.CL2022

Unified Speech-Text Pre-training for Speech Translation and Recognition

Yun Tang, Hongyu Gong, Ning Dong +8

We describe a method to jointly pre-train speech and text in an encoder-decoder modeling framework for speech translation and recognition. The proposed method incorporates four sel…

cs.CL20223 cited

IDPG: An Instance-Dependent Prompt Generation Method

Zhuofeng Wu, Sinong Wang, Jiatao Gu +4

Prompt tuning is a new, efficient NLP transfer learning paradigm that adds a task-specific prompt in each input instance during the model training stage. It freezes the pre-trained…