2 papers
cs.LG2026
DBLP: Phase-Aware Bounded-Loss Transport for Burst-Resilient Distributed ML Training
Zechen Ma, Zixi Qu, Jinyan Yi +2
Distributed machine learning (ML) training has become a necessity with the prevalence of billion to trillion-parameter-scale models. While prior work has improved training efficien…
cs.LG2026
The Geometric Price of Discrete Logic: Context-driven Manifold Dynamics of Number Representations
Long Zhang, Dai-jun Lin, Wei-neng Chen
Large language models (LLMs) generalize smoothly across continuous semantic spaces, yet strict logical reasoning demands the formation of discrete decision boundaries. Prevailing t…