2 papers
cs.AR2026
Hardware-Efficient Softmax and Layer Normalization with Guaranteed Normalization for Edge Devices
Dawon Choi, Hana Kim, Ji-Hoon Kim
In Transformer models, non-GEMM (non-General Matrix Multiplication) operations -- especially Softmax and Layer Normalization (LayerNorm) -- often dominate hardware cost due to thei…
cs.AR2025
Ternary-Input Binary-Weight CNN Accelerator Design for Miniature Object Classification System with Query-Driven Spatial DVS
Yuyang Li, Swasthik Muloor, Jack Laudati +5
Miniature imaging systems are essential for space-constrained applications but are limited by memory and power constraints. While machine learning can reduce data size by extractin…