5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2026
Flow-based Policy With Distributional Reinforcement Learning in Trajectory Optimization
Ruijie Hao, Longfei Zhang, Yang Dai +3
Reinforcement Learning (RL) has proven highly effective in addressing complex control and decision-making tasks. However, in most traditional RL algorithms, the policy is typically…
cs.LG2021★ 5 cited
PTR-PPO: Proximal Policy Optimization with Prioritized Trajectory Replay
Xingxing Liang, Yang Ma, Yanghe Feng +1
On-policy deep reinforcement learning algorithms have low data utilization and require significant experience for policy improvement. This paper proposes a proximal policy optimiza…