1 citations · 2 across the 3 of their papers we have counts for
3 papers
eess.AS2023★ 1 cited
TorchAudio 2.1: Advancing speech recognition, self-supervised learning, and audio processing components for PyTorch
Jeff Hwang, Moto Hira, Caroline Chen +21
TorchAudio is an open-source audio and speech processing library built for PyTorch. It aims to accelerate the research and development of audio and speech technologies by providing…
eess.AS2023★ 1 cited
Multi-Head State Space Model for Speech Recognition
Yassir Fathullah, Chunyang Wu, Yuan Shangguan +8
State space models (SSMs) have recently shown promising results on small-scale sequence and language modelling tasks, rivalling and outperforming many attention-based approaches. I…
eess.AS2022
Learning a Dual-Mode Speech Recognition Model via Self-Pruning
Chunxi Liu, Yuan Shangguan, Haichuan Yang +3
There is growing interest in unifying the streaming and full-context automatic speech recognition (ASR) networks into a single end-to-end ASR model to simplify the model training a…