Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning
Shuyu Wei, Jian Sun, Delai Qiu +6
Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing methods often either increa…
cs.CL2025
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
MiniMax, :, Aili Chen +125
We introduce MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. MiniMax-M1 is powered by a hybrid Mixture-of-Experts (MoE) architecture combin…