Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
AutoMem: A Text-Gradient Recursive Self-Improvement Framework for Automated Memory Architectures Search
Lin Du, Jie Zhou, Yuxuan Cai +6
Long-term memory is increasingly central to LLM agents, yet memory design remains a highly coupled architecture problem: what to encode, how to store it, how to retrieve it, and ho…
cs.CL2025
MiniMax-M1: Scaling Test-Time Compute Efficiently with Lightning Attention
MiniMax, :, Aili Chen +125
We introduce MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model. MiniMax-M1 is powered by a hybrid Mixture-of-Experts (MoE) architecture combin…