1 paper · 1 filter
Joy Jia Yin Lim, Xin Huang, Hao Peng +5
Large language model (LLM) agents can now post-train an LLM end-to-end. They can write code, launch training, evaluate checkpoints, and improve downstream performance, raising the…