1 paper
Cheolseung Baek, Dhammiko Arya, Eunki Kim +40
We introduce A.X K2, a 688B-parameter Mixture-of-Experts (MoE) language model trained from scratch as a high-performance foundation for \emph{agentic} applications. Trained on appr…