4 citations · 6 across the 7 of their papers we have counts for
6 papers · 1 filter
Generative De-Quantization for Neural Speech Codec via Latent Diffusion
Haici Yang, Inseon Jang, Minje Kim
In low-bitrate speech coding, end-to-end speech coding networks aim to learn compact yet expressive features and a powerful decoder in a single network. A challenging problem as su…
Neural Feature Predictor and Discriminative Residual Coding for Low-Bitrate Speech Coding
Haici Yang, Wootaek Lim, Minje Kim
Low and ultra-low-bitrate neural speech coding achieves unprecedented coding gain by generating speech signals from compact speech features. This paper introduces additional coding…
Upmixing via style transfer: a variational autoencoder for disentangling spatial images and musical content
Haici Yang, Sanna Wager, Spencer Russell +3
In the stereo-to-multichannel upmixing problem for music, one of the main tasks is to set the directionality of the instrument sources in the multichannel rendering results. In thi…
Don't Separate, Learn to Remix: End-to-End Neural Remixing with Joint Optimization
Haici Yang, Shivani Firodiya, Nicholas J. Bryan +1
The task of manipulating the level and/or effects of individual instruments to recompose a mixture of recordings, or remixing, is common across a variety of applications such as mu…
Source-Aware Neural Speech Coding for Noisy Speech Compression
Haici Yang, Kai Zhen, Seungkwon Beack +1
This paper introduces a novel neural network-based speech coding system that can process noisy speech effectively. The proposed source-aware neural audio coding (SANAC) system harm…
Boosted Locality Sensitive Hashing: Discriminative Binary Codes for Source Separation
Sunwoo Kim, Haici Yang, Minje Kim
Speech enhancement tasks have seen significant improvements with the advance of deep learning technology, but with the cost of increased computational complexity. In this study, we…