4 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.SD2022★ 4 cited
Leveraging Modality-specific Representations for Audio-visual Speech Recognition via Reinforcement Learning
Chen Chen, Yuchen Hu, Qiang Zhang +3
Audio-visual speech recognition (AVSR) has gained remarkable success for ameliorating the noise-robustness of speech recognition. Mainstream methods focus on fusing audio and visua…
cs.LG2022★ 3 cited
Unleashing the Power of Transformer for Graphs
Lingbing Guo, Qiang Zhang, Huajun Chen
Despite recent successes in natural language processing and computer vision, Transformer suffers from the scalability problem when dealing with graphs. The computational complexity…