4 citations · 8 across the 10 of their papers we have counts for
4 papers · 1 filter
Natural Language Reinforcement Learning
Xidong Feng, Bo Liu, Yan Song +7
Artificial intelligence progresses towards the "Era of Experience," where agents are expected to learn from continuous, grounded interaction. We argue that traditional Reinforcemen…
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
Jun Wang, Meng Fang, Ziyu Wan +10
In this technical report, we introduce OpenR, an open-source framework designed to integrate key components for enhancing the reasoning capabilities of large language models (LLMs)…
Reinforcing Language Agents via Policy Optimization with Action Decomposition
Muning Wen, Ziyu Wan, Weinan Zhang +2
Language models as intelligent agents push the boundaries of sequential decision-making agents but struggle with limited knowledge of environmental dynamics and exponentially huge…
Natural Language Reinforcement Learning
Xidong Feng, Ziyu Wan, Mengyue Yang +5
Reinforcement Learning (RL) has shown remarkable abilities in learning policies for decision-making tasks. However, RL is often hindered by issues such as low sample efficiency, la…