Showing cs.DCShow all
3 papers · 1 filter
cs.DC2026
Cross-Platform Fused MoE Dispatch in Triton: Portable Expert Routing Without CUDA
Subhadip Mitra
Mixture-of-Experts (MoE) architectures power the majority of frontier large language models, but their inference is bottlenecked by irregular memory access patterns and expert rout…
cs.DC2026
Constraint-Aware Execution Planning for Hybrid Space-Ground Compute Workloads
Subhadip Mitra
Low Earth orbit (LEO) satellites increasingly carry compute hardware capable of on-board processing, yet each satellite generates roughly two orders of magnitude more data than it…
cs.DC2026
Spark-LLM-Eval: A Distributed Framework for Statistically Rigorous Large Language Model Evaluation
Subhadip Mitra
Evaluating large language models at scale remains a practical bottleneck for many organizations. While existing evaluation frameworks work well for thousands of examples, they stru…