1 paper
Chen Xu, Yuhao Zhang, Chengbo Jiao +7
While Transformer has become the de-facto standard for speech, modeling upon the fine-grained frame-level features remains an open challenge of capturing long-distance dependencies…