5 papers
QS4D: Quantization-aware training for efficient hardware deployment of structured state-space sequential models
Sebastian Siegel, Ming-Jay Yang, Younes Bouhadjar +3
Structured State Space models (SSM) have recently emerged as a new class of deep learning models, particularly well-suited for processing long sequences. Their constant memory foot…
Real-time raw signal genomic analysis using fully integrated memristor hardware
Peiyi He, Shengbo Wang, Ruibin Mao +6
Advances in third-generation sequencing have enabled portable and real-time genomic sequencing, but real-time data processing remains a bottleneck, hampering on-site genomic analys…
IMSSA: Deploying modern state-space models on memristive in-memory compute hardware
Sebastian Siegel, Ming-Jay Yang, John-Paul Strachan
Processing long temporal sequences is a key challenge in deep learning. In recent years, Transformers have become state-of-the-art for this task, but suffer from excessive memory r…
Analog In-Memory Computing Attention Mechanism for Fast and Energy-Efficient Large Language Models
Nathan Leroux, Paul-Philipp Manea, Chirag Sudarshan +4
Transformer networks, driven by self-attention, are central to Large Language Models. In generative Transformers, self-attention uses cache memory to store token projections, avoid…
The Ouroboros of Memristors: Neural Networks Facilitating Memristor Programming
Zhenming Yu, Ming-Jay Yang, Jan Finkbeiner +3
Memristive devices hold promise to improve the scale and efficiency of machine learning and neuromorphic hardware, thanks to their compact size, low power consumption, and the abil…