1 citations · 1 across the 14 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Kimi K2: Open Agentic Intelligence
Kimi Team, Yifan Bai, Yiping Bao +195
We introduce Kimi K2, a Mixture-of-Experts (MoE) large language model with 32 billion activated parameters and 1 trillion total parameters. We propose the MuonClip optimizer, which…
cs.LG2021★ 1 cited
Learning without Knowing: Unobserved Context in Continuous Transfer Reinforcement Learning
Chenyu Liu, Yan Zhang, Yi Shen +1
In this paper, we consider a transfer Reinforcement Learning (RL) problem in continuous state and action spaces, under unobserved contextual information. For example, the context c…