3 papers
cs.AR2025
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-bit LLMs
Jinendra Malekar, Peyton Chandarana, Md Hasibul Amin +2
In this paper, we propose PIM-LLM, a hybrid architecture developed to accelerate 1-bit large language models (LLMs). PIM-LLM leverages analog processing-in-memory (PIM) architectur…
cs.NE2024
Towards Efficient Deployment of Hybrid SNNs on Neuromorphic and Edge AI Hardware
James Seekings, Peyton Chandarana, Mahsa Ardakani +2
This paper explores the synergistic potential of neuromorphic and edge computing to create a versatile machine learning (ML) system tailored for processing data captured by dynamic…
cs.AR2024
Flex-TPU: A Flexible TPU with Runtime Reconfigurable Dataflow Architecture
Mohammed Elbtity, Peyton Chandarana, Ramtin Zand
Tensor processing units (TPUs) are one of the most well-known machine learning (ML) accelerators utilized at large scale in data centers as well as in tiny ML applications. TPUs of…