Showing eess.ASShow all
2 papers · 1 filter
eess.AS2023
TranUSR: Phoneme-to-word Transcoder Based Unified Speech Representation Learning for Cross-lingual Speech Recognition
Hongfei Xue, Qijie Shao, Peikun Chen +3
UniSpeech has achieved superior performance in cross-lingual automatic speech recognition (ASR) by explicitly aligning latent representations to phoneme units using multi-task self…
eess.AS2023
The NPU-ASLP System for Audio-Visual Speech Recognition in MISP 2022 Challenge
Pengcheng Guo, He Wang, Bingshen Mu +2
This paper describes our NPU-ASLP system for the Audio-Visual Diarization and Recognition (AVDR) task in the Multi-modal Information based Speech Processing (MISP) 2022 Challenge.…