32 citations · 72 across the 9 of their papers we have counts for
4 papers · 1 filter
HGQ: High Granularity Quantization for Real-time Neural Networks on FPGAs
Chang Sun, Zhiqiang Que, Thea K. Årrestad +4
Neural networks with sub-microsecond inference latency are required by many critical applications. Targeting such applications deployed on FPGAs, we present High Granularity Quanti…
Accelerating Recurrent Neural Networks for Gravitational Wave Experiments
Zhiqiang Que, Erwei Wang, Umar Marikar +10
This paper presents novel reconfigurable architectures for reducing the latency of recurrent neural networks (RNNs) that are used for detecting gravitational waves. Gravitational i…
hls4ml: An Open-Source Codesign Workflow to Empower Scientific Low-Power Machine Learning Devices
Farah Fahim, Benjamin Hawks, Christian Herwig +27
Accessible machine learning algorithms, software, and diagnostic tools for energy-efficient devices and systems are extremely valuable across a broad range of application domains.…
Fast convolutional neural networks on FPGAs with hls4ml
Thea Aarrestad, Vladimir Loncar, Nicolò Ghielmetti +17
We introduce an automated tool for deploying ultra low-latency, low-power deep neural networks with convolutional layers on FPGAs. By extending the hls4ml library, we demonstrate a…