1.6k citations
- Ministry of Industry and Information TechnologyCN18 papers
- Tsinghua UniversityCN18 papers
- Xi'an Jiaotong UniversityCN18 papers
- Xidian UniversityCN18 papers
- Chinese Academy of SciencesCN17 papers
- Australian National UniversityAU11 papers
- Moscow Institute of Physics and TechnologyRU11 papers
- Nanyang Technological UniversitySG11 papers
- University of Chinese Academy of SciencesCN11 papers
- Beijing Normal UniversityCN9 papers
- Fudan UniversityCN9 papers
- Nankai UniversityCN9 papers
4 papers · 2 filters
Conformer-based End-to-end Speech Recognition With Rotary Position Embedding
Shengqiang Li, Menglong Xu, Xiao-Lei Zhang
Transformer-based end-to-end speech recognition models have received considerable attention in recent years due to their high training speed and ability to model a long-range globa…
Audio Description from Image by Modal Translation Network
Hailong Ning, Xiangtao Zheng, Yuan Yuan +1
Audio is the main form for the visually impaired to obtain information. In reality, all kinds of visual data always exist, but audio data does not exist in many cases. In order to…
An Asynchronous WFST-Based Decoder For Automatic Speech Recognition
Hang Lv, Zhehuai Chen, Hainan Xu +3
We introduce asynchronous dynamic decoder, which adopts an efficient A* algorithm to incorporate big language models in the one-pass decoding for large vocabulary continuous speech…
The Accented English Speech Recognition Challenge 2020: Open Datasets, Tracks, Baselines, Results and Methods
Xian Shi, Fan Yu, Yizhou Lu +5
The variety of accents has posed a big challenge to speech recognition. The Accented English Speech Recognition Challenge (AESRC2020) is designed for providing a common testbed and…