2 papers
cs.AR2026
Arithmetic Packing on Wide Integer Datapaths in DSP Primitives of Modern FPGA Devices
Titus Bornträger, Shane Fleming, Philipp Holzinger +3
Deep Neural Networks increasingly employ low-precision quantization to reduce computational requirements. While FPGAs are well suited for workloads with heterogeneous precisions, t…
cs.AR2025
Multiplier-free In-Memory Vector-Matrix Multiplication Using Distributed Arithmetic
Felix Zeller, John Reuben, Dietmar Fey
Vector-Matrix Multiplication (VMM) is the fundamental and frequently required computation in inference of Neural Networks (NN). Due to the large data movement required during infer…