3 papers
cs.CV2024
Memory-Efficient Vision Transformers: An Activation-Aware Mixed-Rank Compression Strategy
Seyedarmin Azizi, Mahdi Nazemi, Massoud Pedram
As Vision Transformers (ViTs) increasingly set new benchmarks in computer vision, their practical deployment on inference engines is often hindered by their significant memory band…
cs.AR2023
Algorithms and Hardware for Efficient Processing of Logic-based Neural Networks
Jingkai Hong, Arash Fayyazi, Amirhossein Esmaili +2
Recent efforts to improve the performance of neural network (NN) accelerators that meet today's application requirements have given rise to a new trend of logic-based NN inference…
cs.AR2022
Efficient Compilation and Mapping of Fixed Function Combinational Logic onto Digital Signal Processors Targeting Neural Network Inference and Utilizing High-level Synthesis
Soheil Nazar Shahsavani, Arash Fayyazi, Mahdi Nazemi +1
Recent efforts for improving the performance of neural network (NN) accelerators that meet today's application requirements have given rise to a new trend of logic-based NN inferen…