8 citations · 21 across the 9 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2025
Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning
Xinsong Feng, Zihan Yu, Yanhai Xiong +1
Reinforcement learning (RL) has emerged as a promising tool for combinatorial optimization (CO) problems due to its ability to learn fast, effective, and generalizable solutions. N…
cs.AI2024
Focused ReAct: Improving ReAct through Reiterate and Early Stop
Shuoqiu Li, Han Xu, Haipeng Chen
Large language models (LLMs) have significantly improved their reasoning and decision-making capabilities, as seen in methods like ReAct. However, despite its effectiveness in tack…
cs.AI2020★ 8 cited
Learning Behaviors with Uncertain Human Feedback
Xu He, Haipeng Chen, Bo An
Human feedback is widely used to train agents in many domains. However, previous works rarely consider the uncertainty when humans provide feedback, especially in cases that the op…