31 citations · 49 across the 4 of their papers we have counts for
1 paper · 1 filter
Yang Zhang, Krishna C. Puvvada, Vitaly Lavrukhin +1
We propose CONF-TSASR, a non-autoregressive end-to-end time-frequency domain architecture for single-channel target-speaker automatic speech recognition (TS-ASR). The model consist…