3 papers
cs.DC2019
Towards a Uniform Architecture for the Efficient Implementation of 2D and 3D Deconvolutional Neural Networks on FPGAs
Deguang Wang, Junzhong Shen, Mei Wen +1
Three-dimensional deconvolution is widely used in many computer vision applications. However, most previous works have only focused on accelerating 2D deconvolutional neural networ…
cs.NI2019
Metaflow: A DAG-Based Network Abstraction for Distributed Applications
Jiawei Fei, Yang Shi, Qun Huang +1
In the past decade, increasingly network scheduling techniques have been proposed to boost the distributed application performance. Flow-level metrics, such as flow completion time…
cs.AR2018
Towards a Multi-array Architecture for Accelerating Large-scale Matrix Multiplication on FPGAs
Junzhong Shen, Yuran Qiao, You Huang +2
Large-scale floating-point matrix multiplication is a fundamental kernel in many scientific and engineering applications. Most existing work only focus on accelerating matrix multi…