78 citations · 127 across the 9 of their papers we have counts for
3 papers · 1 filter
Improved disentangled speech representations using contrastive learning in factorized hierarchical variational autoencoder
Yuying Xie, Thomas Arildsen, Zheng-Hua Tan
Leveraging the fact that speaker identity and content vary on different time scales, \acrlong{fhvae} (\acrshort{fhvae}) uses different latent variables to symbolize these two attri…
Disentangled Speech Representation Learning Based on Factorized Hierarchical Variational Autoencoder with Self-Supervised Objective
Yuying Xie, Thomas Arildsen, Zheng-Hua Tan
Disentangled representation learning aims to extract explanatory features or factors and retain salient information. Factorized hierarchical variational autoencoder (FHVAE) present…
Complex Recurrent Variational Autoencoder with Application to Speech Enhancement
Yuying Xie, Thomas Arildsen, Zheng-Hua Tan
As an extension of variational autoencoder (VAE), complex VAE uses complex Gaussian distributions to model latent variables and data. This work proposes a complex recurrent VAE fra…