6 citations · 6 across the 1 of their papers we have counts for
3 papers
cs.MM2022★ 6 cited
Predict-and-Update Network: Audio-Visual Speech Recognition Inspired by Human Speech Perception
Jiadong Wang, Xinyuan Qian, Haizhou Li
Audio and visual signals complement each other in human speech perception, so do they in speech recognition. The visual hint is less evident than the acoustic hint, but more robust…
cs.SD2021
Multi-target DoA Estimation with an Audio-visual Fusion Mechanism
Xinyuan Qian, Maulik Madhavi, Zexu Pan +2
Most of the prior studies in the spatial \ac{DoA} domain focus on a single modality. However, humans use auditory and visual senses to detect the presence of sound sources. With th…
cs.NE2020
Rectified Linear Postsynaptic Potential Function for Backpropagation in Deep Spiking Neural Networks
Malu Zhang, Jiadong Wang, Burin Amornpaisannon +8
Spiking Neural Networks (SNNs) use spatio-temporal spike patterns to represent and transmit information, which is not only biologically realistic but also suitable for ultra-low-po…