activity
20232025
most citedFAMOUS: Flexible Accelerator for the Attention Mechanism of Transformer on UltraScale+ FPGAs

5 citations · 5 across the 5 of their papers we have counts for

collaborators
Showing cs.ARShow all

5 papers · 1 filter

cs.AR2025

N-TORC: Native Tensor Optimizer for Real-time Constraints

Suyash Vardhan Singh, Iftakhar Ahmad, David Andrews +3

Compared to overlay-based tensor architectures like VTA or Gemmini, compilers that directly translate machine learning models into a dataflow architecture as HLS code, such as HLS4…

cs.AR2024

A Runtime-Adaptive Transformer Neural Network Accelerator on FPGAs

Ehsan Kabir, Jason D. Bakos, David Andrews +1

Transformer neural networks (TNN) excel in natural language processing (NLP), machine translation, and computer vision (CV) without relying on recurrent or convolutional layers. Ho…

cs.AR2024

ProTEA: Programmable Transformer Encoder Acceleration on FPGA

Ehsan Kabir, Jason D. Bakos, David Andrews +1

Transformer neural networks (TNN) have been widely utilized on a diverse range of applications, including natural language processing (NLP), machine translation, and computer visio…

cs.AR20245 cited

FAMOUS: Flexible Accelerator for the Attention Mechanism of Transformer on UltraScale+ FPGAs

Ehsan Kabir, Md. Arafat Kabir, Austin R. J. Downey +3

Transformer neural networks (TNNs) are being applied across a widening range of application domains, including natural language processing (NLP), machine translation, and computer…

cs.AR2023

Accelerating LSTM-based High-Rate Dynamic System Models

Ehsan Kabir, Daniel Coble, Joud N. Satme +4

In this paper, we evaluate the use of a trained Long Short-Term Memory (LSTM) network as a surrogate for a Euler-Bernoulli beam model, and then we describe and characterize an FPGA…