3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.AR2025★ 1 cited
Pushing up to the Limit of Memory Bandwidth and Capacity Utilization for Efficient LLM Decoding on Embedded FPGA
Jindong Li, Tenglong Li, Guobin Shen +3
The extremely high computational and storage demands of large language models have excluded most edge devices, which were widely used for efficient machine learning, from being via…
cs.AR2024
Revealing Untapped DSP Optimization Potentials for FPGA-Based Systolic Matrix Engines
Jindong Li, Tenglong Li, Guobin Shen +3
Systolic architectures are widely embraced by neural network accelerators for their superior performance in highly parallelized computation. The DSP48E2s serve as dedicated arithme…
math.NA2022★ 3 cited
SMS: Spiking Marching Scheme for Efficient Long Time Integration of Differential Equations
Qian Zhang, Adar Kahana, George Em Karniadakis +1
We propose a Spiking Neural Network (SNN)-based explicit numerical scheme for long time integration of time-dependent Ordinary and Partial Differential Equations (ODEs, PDEs). The…