Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
History Rhymes: Accelerating LLM Reinforcement Learning with RhymeRL
Jingkai He, Tianjian Li, Erhu Feng +5
With the rapid advancement of large language models (LLMs), reinforcement learning (RL) has emerged as a pivotal methodology for enhancing the reasoning capabilities of LLMs. Unlik…
cs.LG2025
Get Experience from Practice: LLM Agents with Record & Replay
Erhu Feng, Wenbo Zhou, Zibin Liu +8
AI agents, empowered by Large Language Models (LLMs) and communication protocols such as MCP and A2A, have rapidly evolved from simple chatbots to autonomous entities capable of ex…