Showing cs.DCShow all
2 papers · 1 filter
cs.DC2024
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training
Kun Wu, Jeongmin Brian Park, Xiaofan Zhang +5
The growth rate of the GPU memory capacity has not been able to keep up with that of the size of large language models (LLMs), hindering the model training process. In particular,…
cs.DC2018
ASAP: Accelerated Short-Read Alignment on Programmable Hardware
Subho S. Banerjee, Mohamed El-Hadedy, Jong Bin Lim +4
The proliferation of high-throughput sequencing machines ensures rapid generation of up to billions of short nucleotide fragments in a short period of time. This massive amount of…