16 citations · 16 across the 2 of their papers we have counts for
2 papers
eess.AS2022
Progressive Multi-Scale Self-Supervised Learning for Speech Recognition
Genshun Wan, Tan Liu, Hang Chen +3
Self-supervised learning (SSL) models have achieved considerable improvements in automatic speech recognition (ASR). In addition, ASR performance could be further improved if the m…
cs.CL2022★ 16 cited
A Deep Neural Framework for Image Caption Generation Using GRU-Based Attention Mechanism
Rashid Khan, M Shujah Islam, Khadija Kanwal +3
Image captioning is a fast-growing research field of computer vision and natural language processing that involves creating text explanations for images. This study aims to develop…