7 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.DC2026
DALI: A Workload-Aware Offloading Framework for Efficient MoE Inference on Local PCs
Zeyu Zhu, Gang Li, Peisong Wang +5
Mixture of Experts (MoE) architectures significantly enhance the capacity of LLMs without proportional increases in computation, but at the cost of a vast parameter size. Offloadin…
cs.AR2025
GCC: A 3DGS Inference Architecture with Gaussian-Wise and Cross-Stage Conditional Processing
Minnan Pei, Gang Li, Junwen Si +6
3D Gaussian Splatting (3DGS) has emerged as a leading neural rendering technique for high-fidelity view synthesis, prompting the development of dedicated 3DGS accelerators for reso…
cs.LG2024★ 7 cited
FastGL: A GPU-Efficient Framework for Accelerating Sampling-Based GNN Training at Large Scale
Zeyu Zhu, Peisong Wang, Qinghao Hu +3
Graph Neural Networks (GNNs) have shown great superiority on non-Euclidean graph data, achieving ground-breaking performance on various graph-related tasks. As a practical solution…