activity
20212023
most citedEVC: Towards Real-Time Neural Image Compression with Mask Decay

24 citations · 45 across the 17 of their papers we have counts for

collaborators
Showing cs.SDShow all

6 papers · 1 filter

cs.SD20231 cited

DasFormer: Deep Alternating Spectrogram Transformer for Multi/Single-Channel Speech Separation

Shuo Wang, Xiangyu Kong, Xiulian Peng +3

For the task of speech separation, previous study usually treats multi-channel and single-channel scenarios as two research tracks with specialized solutions developed respectively…

cs.SD20231 cited

Contrast-PLC: Contrastive Learning for Packet Loss Concealment

Huaying Xue, Xiulian Peng, Yan Lu

Packet loss concealment (PLC) is challenging in concealing missing contents both plausibly and naturally when there are only limited available context to use. Recently deep-learnin…

cs.SD2023

Improving Speech Enhancement via Event-based Query

Yifei Xin, Xiulian Peng, Yan Lu

Existing deep learning based speech enhancement (SE) methods either use blind end-to-end training or explicitly incorporate speaker embedding or phonetic information into the SE ne…

cs.SD2022

Cross-Scale Vector Quantization for Scalable Neural Speech Coding

Xue Jiang, Xiulian Peng, Huaying Xue +2

Bitrate scalability is a desirable feature for audio coding in real-time communications. Existing neural audio codecs usually enforce a specific bitrate during training, so differe…

cs.SD2022

Multi-Modal Multi-Correlation Learning for Audio-Visual Speech Separation

Xiaoyu Wang, Xiangyu Kong, Xiulian Peng +1

In this paper we propose a multi-modal multi-correlation learning framework targeting at the task of audio-visual speech separation. Although previous efforts have been extensively…

cs.SD2022

Towards Error-Resilient Neural Speech Coding

Huaying Xue, Xiulian Peng, Xue Jiang +1

Neural audio coding has shown very promising results recently in the literature to largely outperform traditional codecs but limited attention has been paid on its error resilience…