2 papers
cs.AI2025
CIMNAS: A Joint Framework for Compute-In-Memory-Aware Neural Architecture Search
Olga Krestinskaya, Mohammed E. Fouda, Ahmed Eltawil +1
To maximize hardware efficiency and performance accuracy in Compute-In-Memory (CIM)-based neural network accelerators for Artificial Intelligence (AI) applications, co-optimizing b…
cs.AR2024
SoftmAP: Software-Hardware Co-design for Integer-Only Softmax on Associative Processors
Mariam Rakka, Jinhao Li, Guohao Dai +3
Recent research efforts focus on reducing the computational and memory overheads of Large Language Models (LLMs) to make them feasible on resource-constrained devices. Despite adva…