2 papers
cs.RO2021
Reward Shaping with Subgoals for Social Navigation
Takato Okudo, Seiji Yamada
Social navigation has been gaining attentions with the growth in machine intelligence. Since reinforcement learning can select an action in the prediction phase at a low computatio…
cs.LG2021
Reward Shaping with Dynamic Trajectory Aggregation
Takato Okudo, Seiji Yamada
Reinforcement learning, which acquires a policy maximizing long-term rewards, has been actively studied. Unfortunately, this learning type is too slow and difficult to use in pract…