15 citations · 15 across the 3 of their papers we have counts for
3 papers
SoftmAP: Software-Hardware Co-design for Integer-Only Softmax on Associative Processors
Mariam Rakka, Jinhao Li, Guohao Dai +3
Recent research efforts focus on reducing the computational and memory overheads of Large Language Models (LLMs) to make them feasible on resource-constrained devices. Despite adva…
BF-IMNA: A Bit Fluid In-Memory Neural Architecture for Neural Network Acceleration
Mariam Rakka, Rachid Karami, Ahmed M. Eltawil +2
Mixed-precision quantization works Neural Networks (NNs) are gaining traction for their efficient realization on the hardware leading to higher throughput and lower energy. In-Memo…
Mixed-Precision Neural Networks: A Survey
Mariam Rakka, Mohammed E. Fouda, Pramod Khargonekar +1
Mixed-precision Deep Neural Networks achieve the energy efficiency and throughput needed for hardware deployment, particularly when the resources are limited, without sacrificing a…