2 papers
cs.LG2024
MoNTA: Accelerating Mixture-of-Experts Training with Network-Traffc-Aware Parallel Optimization
Jingming Guo, Yan Liu, Yu Meng +4
The Mixture of Experts (MoE) is an advanced model architecture in the industry that combines multiple specialized expert models from various domains into a single supermodel. This…
quant-ph2024
Basis-independent quantum coherence and its distribution under relativistic motion
Ming-Ming Du, Hong-Wei Li, Zhen Tao +5
Recent studies have increasingly focused on the effect of relativistic motion on quantum coherence. Prior research predominantly examined the influence of relative motion on basis-…