3 papers
cs.DC2026
GEM: GPU-Variability-Aware Expert to GPU Mapping for MoE Systems
Sourish Wawdhane, Avinash Kumar, Poulami Das
Mixture-of-Expert (MoE) models enable efficient inference by employing smaller experts and activating only a subset of them per token. MoE serving engines distribute experts across…
quant-ph2026
Are LLMs Good For Quantum Software, Architecture, and System Design?
Sourish Wawdhane, Poulami Das
Quantum computers promise massive computational speedup for problems in many critical domains, such as physics, chemistry, cryptanalysis, healthcare, etc. However, despite decades…
quant-ph2025
Ensemble-IR: Concise Representation for Quantum Ensemble Programs
Sourish Wawdhane, Sashwat Anagolum, Poulami Das +1
Emerging quantum applications such as error mitigation, system characterization, and hybrid protocols often require running large families of related quantum circuits. Existing int…