2 papers
cs.AR2026
CIMple: Standard-cell SRAM-based CIM with LUT-based split softmax for attention acceleration
Bas Ahn, Xingjian Tao, Manil Dev Gomony +2
Large Language Models (LLMs) such as LLaMA and DeepSeek, are built on transformer architectures, which have become a standard model for achieving state-of-the-art performance in na…
cs.AR2025
A Unified Framework for Mapping and Synthesis of Approximate R-Blocks CGRAs
Georgios Alexandris, Panagiotis Chaidos, Alexis Maras +5
The ever-increasing complexity and operational diversity of modern Neural Networks (NNs) have caused the need for low-power and, at the same time, high-performance edge devices for…