output
20152025
most citedSimple and Deep Graph Convolutional Networks

402 citations

Showing cs.SDShow all

5 papers · 1 filter

cs.SD2021

BeamTransformer: Microphone Array-based Overlapping Speech Detection

Siqi Zheng, Shiliang Zhang, Weilong Huang +5

We propose BeamTransformer, an efficient architecture to leverage beamformer's edge in spatial filtering and transformer's capability in context sequence modeling. BeamTransformer…

cs.SD202117 cited

A Real-time Speaker Diarization System Based on Spatial Spectrum

Siqi Zheng, Weilong Huang, Xianliang Wang +3

In this paper we describe a speaker diarization system that enables localization and identification of all speakers present in a conversation or meeting. We propose a novel systema…

cs.SD20211 cited

Weighted Recursive Least Square Filter and Neural Network based Residual Echo Suppression for the AEC-Challenge

Ziteng Wang, Yueyue Na, Zhang Liu +2

This paper presents a real-time Acoustic Echo Cancellation (AEC) algorithm submitted to the AEC-Challenge. The algorithm consists of three modules: Generalized Cross-Correlation wi…

cs.SD20211 cited

Towards Natural and Controllable Cross-Lingual Voice Conversion Based on Neural TTS Model and Phonetic Posteriorgram

Shengkui Zhao, Hao Wang, Trung Hieu Nguyen +1

Cross-lingual voice conversion (VC) is an important and challenging problem due to significant mismatches of the phonetic set and the speech prosody of different languages. In this…

cs.SD20208 cited

INT8 Winograd Acceleration for Conv1D Equipped ASR Models Deployed on Mobile Devices

Yiwu Yao, Yuchao Li, Chengyu Wang +8

The intensive computation of Automatic Speech Recognition (ASR) models obstructs them from being deployed on mobile devices. In this paper, we present a novel quantized Winograd op…