1 citations · 1 across the 3 of their papers we have counts for
3 papers
LLMPi: Optimizing LLMs for High-Throughput on Raspberry Pi
Mahsa Ardakani, Jinendra Malekar, Ramtin Zand
Deploying Large Language Models (LLMs) on resource-constrained edge devices like the Raspberry Pi presents challenges in computational efficiency, power consumption, and response l…
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
Jinendra Malekar, Peyton Chandarana, Md Hasibul Amin +2
In this paper, we propose PIM-LLM, a hybrid architecture developed to accelerate 1-bit large language models (LLMs). PIM-LLM leverages analog processing-in-memory (PIM) architectur…
Matmul or No Matmul in the Era of 1-bit LLMs
Jinendra Malekar, Mohammed E. Elbtity, Ramtin Zand
The advent of 1-bit large language models (LLMs) has attracted considerable attention and opened up new research opportunities. However, 1-bit LLMs only improve a fraction of model…