16 citations · 16 across the 2 of their papers we have counts for
4 papers
A Deep Learning Inference Scheme Based on Pipelined Matrix Multiplication Acceleration Design and Non-uniform Quantization
Yuyang Zhang, Dik Hin Leung, Min Guo +6
Matrix multiplication is the bedrock in Deep Learning inference application. When it comes to hardware acceleration on edge computing devices, matrix multiplication often takes up…
ViP: Virtual Pooling for Accelerating CNN-based Image Classification and Object Detection
Zhuo Chen, Jiyuan Zhang, Ruizhou Ding +1
In recent years, Convolutional Neural Networks (CNNs) have shown superior capability in visual learning tasks. While accuracy-wise CNNs provide unprecedented performance, they are…
Memory-Attended Recurrent Network for Video Captioning
Wenjie Pei, Jiyuan Zhang, Xiangrong Wang +3
Typical techniques for video captioning follow the encoder-decoder framework, which can only focus on one source video being processed. A potential disadvantage of such design is t…
High Performance Zero-Memory Overhead Direct Convolutions
Jiyuan Zhang, Franz Franchetti, Tze Meng Low
The computation of convolution layers in deep neural networks typically rely on high performance routines that trade space for time by using additional memory (either for packing p…