activity
20152022
most citedAUTOVC: Zero-Shot Voice Style Transfer with Only Autoencoder Loss

195 citations · 691 across the 35 of their papers we have counts for

collaborators

59 papers

cs.CL20226 cited

DiffCSE: Difference-based Contrastive Learning for Sentence Embeddings

Yung-Sung Chuang, Rumen Dangovski, Hongyin Luo +7

We propose DiffCSE, an unsupervised contrastive learning framework for learning sentence embeddings. DiffCSE learns sentence embeddings that are sensitive to the difference between…

eess.AS20221 cited

WAVPROMPT: Towards Few-Shot Spoken Language Understanding with Frozen Language Models

Heting Gao, Junrui Ni, Kaizhi Qian +3

Large-scale auto-regressive language models pretrained on massive text have demonstrated their impressive ability to perform new natural language tasks with only a few text example…

cs.LG20223 cited

Adversarial Support Alignment

Shangyuan Tong, Timur Garipov, Yang Zhang +2

We study the problem of aligning the supports of distributions. Compared to the existing work on distribution alignment, support alignment does not require the densities to be matc…

cs.SD2021

On the Interplay Between Sparsity, Naturalness, Intelligibility, and Prosody in Speech Synthesis

Cheng-I Jeff Lai, Erica Cooper, Yang Zhang +8

Are end-to-end text-to-speech (TTS) models over-parametrized? To what extent can these models be pruned, and what happens to their synthesis capabilities? This work serves as a sta…

cs.LG202118 cited

Understanding Interlocking Dynamics of Cooperative Rationalization

Mo Yu, Yang Zhang, Shiyu Chang +1

Selective rationalization explains the prediction of complex neural networks by finding a small subset of the input that is sufficient to predict the neural model output. The selec…

eess.AS202112 cited

Global Rhythm Style Transfer Without Text Transcriptions

Kaizhi Qian, Yang Zhang, Shiyu Chang +4

Prosody plays an important role in characterizing the style of a speaker or an emotion, but most non-parallel voice or emotion style transfer algorithms do not convert any prosody…