An Electro-Photonic System for Accelerating Deep Neural Networks
arXiv:2109.01126 · doi:10.1145/3606949
Abstract
The number of parameters in deep neural networks (DNNs) is scaling at about 5 the rate of Moore's Law. To sustain this growth, photonic computing is a promising avenue, as it enables higher throughput in dominant general matrix-matrix multiplication (GEMM) operations in DNNs than their electrical counterpart. However, purely photonic systems face several challenges including lack of photonic memory and accumulation of noise. In this paper, we present an electro-photonic accelerator, ADEPT, which leverages a photonic computing unit for performing GEMM operations, a vectorized digital electronic ASIC for performing non-GEMM operations, and SRAM arrays for storing DNN parameters and activations. In contrast to prior works in photonic DNN accelerators, we adopt a system-level perspective and show that the gains while large are tempered relative to prior expectations. Our goal is to encourage architects to explore photonic technology in a more pragmatic way considering the system as a whole to understand its general applicability in accelerating today's DNNs. Our evaluation shows that ADEPT can provide, on average, 5.73 higher throughput per Watt compared to the traditional systolic arrays (SAs) in a full-system, and at least 6.8 and better throughput per Watt, compared to state-of-the-art electronic and photonic accelerators, respectively.
References in corpus (15)
- Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
- Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
- Language Models are Few-Shot Learners
- Photonics for artificial intelligence and neuromorphic computing
- 11 TeraFLOPs per second photonic convolutional accelerator for deep learning optical neural networks
- cuDNN: Efficient Primitives for Deep Learning
- Efficient, Compact and Low Loss Thermo-Optic Phase Shifter in Silicon
- Integer Quantization for Deep Learning Inference: Principles and Empirical Evaluation
- Hardware error correction for programmable photonics
- 60dB high-extinction auto-configured Mach--Zehnder interferometer
- Scaling Up Silicon Photonic-based Accelerators: Challenges and Opportunities
- Accurate Self-Configuration of Rectangular Multiport Interferometers
- Low-memory GEMM-based convolution algorithms for deep neural networks
- Stability of Self-Configuring Large Multiport Interferometers
- Optical Convolutional Neural Networks -- Combining Silicon Photonics and Fourier Optics for Computer Vision