Showing cs.DCShow all
3 papers · 1 filter
cs.DC2026
HIERA: Workload-Aware Planning Across Implementation Spaces for GPU Kernel Optimization
Jinghao Wang, Qiqi Gu, Chenpeng Wu +3
High-performance GPU kernels underpin modern deep learning and scientific computing. As workloads become increasingly diverse and GPU hardware evolves rapidly, developing efficient…
cs.DC2026
FWeb3: A Practical Incentive-Aware Federated Learning Framework
Peishen Yan, Shuang Liang, Yang Hua +9
Federated learning (FL) enables collaborative model training over distributed private data. However, sustaining open participation requires incentive mechanisms that compensate con…
cs.DC2026
Do We Need Tensor Cores for Stencil Computations?
Qiqi Gu, Chenpeng Wu, Heng Shi +2
Stencil computation constitutes a cornerstone of scientific computing, serving as a critical kernel in domains ranging from fluid dynamics to weather simulation. While stencil comp…