demonstration retrieval 1generative models 1offline reinforcement learning 1policy generalization 1retrieval-based planning 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making
Lu Guo, Yixiang Shan, Zhengbang Zhu +5
The paper proposes RAD, a method that improves offline reinforcement learning by retrieving high-return states from the dataset and generating sub-trajectories toward these targets…
cs.AI2026
Capability-Aligned Hierarchical Learning for Tool-Augmented LLMs
Haotong Yang, Ting Long, Yi Chang
Tool learning enables LLMs to invoke external tools to accomplish tasks. Prior studies have demonstrated the effectiveness of a hierarchical structure: a high-level policy handles…