5 papers
RETENTION: Resource-Efficient Tree-Based Ensemble Model Acceleration with Content-Addressable Memory
Yi-Chun Liao, Chieh-Lin Tsai, Yuan-Hao Chang +3
Although deep learning has demonstrated remarkable capability in learning from unstructured data, modern tree-based ensemble models remain superior in extracting relevant informati…
Sensitivity-Aware Mixed-Precision Quantization for ReRAM-based Computing-in-Memory
Guan-Cheng Chen, Chieh-Lin Tsai, Pei-Hsuan Tsai +1
Compute-In-Memory (CIM) systems, particularly those utilizing ReRAM and memristive technologies, offer a promising path toward energy-efficient neural network computation. However,…
SARA: A Stall-Aware Memory Allocation Strategy for Mixed-Criticality Systems
Meng-Chia Lee, Wen Sheng Lim, Yuan-Hao Chang +1
The memory capacity in edge devices is often limited due to constraints on cost, size, and power. Consequently, memory competition leads to inevitable page swapping in memory-const…
ReCross: Efficient Embedding Reduction Scheme for In-Memory Computing using ReRAM-Based Crossbar
Yu-Hong Lai, Chieh-Lin Tsai, Wen Sheng Lim +3
Deep learning-based recommendation models (DLRMs) are widely deployed in commercial applications to enhance user experience. However, the large and sparse embedding layers in these…
Search-in-Memory (SiM): Reliable, Versatile, and Efficient Data Matching in SSD's NAND Flash Memory Chip for Data Indexing Acceleration
Yun-Chih Chen, Yuan-Hao Chang, Tei-Wei Kuo
To index the increasing volume of data, modern data indexes are typically stored on SSDs and cached in DRAM. However, searching such an index has resulted in significant I/O traffi…