agentic reinforcement learning 1on-policy distillation 1sample efficiency 1skill extraction 1text-based agents 1
From the 1 of 26 linked papers with an AI index.
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
CaveAgent: Transforming LLMs into Stateful Runtime Operators
Maohao Ran, Zhenglin Wan, Cooper Lin +21
LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that struggle with long-horizon tasks…
cs.AI2025
Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors
Fan Nie, Lan Feng, Haotian Ye +5
Efficiently leveraging of the capabilities of contemporary large language models (LLMs) is increasingly challenging, particularly when direct fine-tuning is expensive and often imp…