1 paper · 1 filter
Lunjun Zhang, Ryan Chen, Bradly C. Stadie
Building agentic systems that can autonomously self-improve from experience is a longstanding goal of AI. Large language models (LLMs) today primarily self-improve via two mechanis…