activity
20222026
most citedBrainTTA: A 35 fJ/op Compiler Programmable Mixed-Precision Transport-Triggered NN SoC

3 citations · 4 across the 8 of their papers we have counts for

collaborators
Showing cs.ARShow all

5 papers · 1 filter

cs.AR2026

Mega: A 22 nm Convolutional Spiking Neural Network Accelerator Achieving 0.375 pJ/SOP for Efficient Edge Vision

Rick Luiken, Manil Dev Gomony, Sander Stuijk

Convolutional Spiking Neural Networks (SNN) offer the potential for highly energy-efficient vision processing by exploiting sparse, event-driven computation. However, existing SNN…

cs.AR2026

CIMple: Standard-cell SRAM-based CIM with LUT-based split softmax for attention acceleration

Bas Ahn, Xingjian Tao, Manil Dev Gomony +2

Large Language Models (LLMs) such as LLaMA and DeepSeek, are built on transformer architectures, which have become a standard model for achieving state-of-the-art performance in na…

cs.AR2025

LinkBo: An Adaptive Single-Wire, Low-Latency, and Fault-Tolerant Communications Interface for Variable-Distance Chip-to-Chip Systems

Bochen Ye, Gustavo Naspolini, Kimmo Salo +1

Cost-effective embedded systems necessitate utilizing the single-wire communication protocol for inter-chip communication, thanks to its reduced pin count in comparison to the mult…

cs.AR2025

A Unified Framework for Mapping and Synthesis of Approximate R-Blocks CGRAs

Georgios Alexandris, Panagiotis Chaidos, Alexis Maras +5

The ever-increasing complexity and operational diversity of modern Neural Networks (NNs) have caused the need for low-power and, at the same time, high-performance edge devices for…

cs.AR20223 cited

BrainTTA: A 35 fJ/op Compiler Programmable Mixed-Precision Transport-Triggered NN SoC

Maarten Molendijk, Floran de Putter, Manil Gomony +2

Recently, accelerators for extremely quantized deep neural network (DNN) inference with operand widths as low as 1-bit have gained popularity due to their ability to largely cut do…