2 papers
eess.AS2022
High Quality Audio Coding with MDCTNet
Grant Davidson, Mark Vinton, Per Ekstrand +3
We propose a neural audio generative model, MDCTNet, operating in the perceptually weighted domain of an adaptive modified discrete cosine transform (MDCT). The architecture of the…
eess.AS2022
Stereo Speech Enhancement Using Custom Mid-Side Signals and Monaural Processing
Aaron Master, Lie Lu, Nathan Swedlow
Speech Enhancement (SE) systems typically operate on monaural input and are used for applications including voice communications and capture cleanup for user generated content. Rec…