3 papers
cs.DC2024
Towards Fast Setup and High Throughput of GPU Serverless Computing
Han Zhao, Weihao Cui, Quan Chen +6
Integrating GPUs into serverless computing platforms is crucial for improving efficiency. However, existing solutions for GPU-enabled serverless computing platforms face two signif…
cs.DC2024
Accelerating Sparse DNNs Based on Tiled GEMM
Cong Guo, Fengchen Xue, Jingwen Leng +5
Network pruning can reduce the computation cost of deep neural network (DNN) models. However, sparse models often produce randomly-distributed weights to maintain accuracy, leading…
cs.DC2023
AdaptGear: Accelerating GNN Training via Adaptive Subgraph-Level Kernels on GPUs
Yangjie Zhou, Yaoxu Song, Jingwen Leng +7
Graph neural networks (GNNs) are powerful tools for exploring and learning from graph structures and features. As such, achieving high-performance execution for GNNs becomes crucia…