3 citations · 5 across the 4 of their papers we have counts for
1 paper · 1 filter
Su Zhang, Yi Ding, Ziquan Wei +1
We propose an audio-visual spatial-temporal deep neural network with: (1) a visual block containing a pretrained 2D-CNN followed by a temporal convolutional network (TCN); (2) an a…