5 citations · 10 across the 5 of their papers we have counts for
6 papers
Text-Based Action-Model Acquisition for Planning
Kebing Jin, Huaixun Chen, Hankz Hankui Zhuo
Although there have been approaches that are capable of learning action models from plan traces, there is no work on learning action models from textual observations, which is perv…
Introduction to The Dynamic Pickup and Delivery Problem Benchmark -- ICAPS 2021 Competition
Jianye Hao, Jiawen Lu, Xijun Li +4
The Dynamic Pickup and Delivery Problem (DPDP) is an essential problem within the logistics domain. So far, research on this problem has mainly focused on using artificial data whi…
Coordinated Proximal Policy Optimization
Zifan Wu, Chao Yu, Deheng Ye +3
We present Coordinated Proximal Policy Optimization (CoPPO), an algorithm that extends the original Proximal Policy Optimization (PPO) to the multi-agent setting. The key idea lies…
Transfer Value Iteration Networks
Junyi Shen, Hankz Hankui Zhuo, Jin Xu +2
Value iteration networks (VINs) have been demonstrated to have a good generalization ability for reinforcement learning tasks across similar domains. However, based on our experime…
Repositioning Bikes with Carrier Vehicles and Bike Trailers in Bike Sharing Systems
Xinghua Zheng, Ming Tang, Hankz Hankui Zhuo +1
Bike Sharing Systems (BSSs) have been adopted in many major cities of the world due to traffic congestion and carbon emissions. Although there have been approaches to exploiting ei…
Federated Deep Reinforcement Learning
Hankz Hankui Zhuo, Wenfeng Feng, Yufeng Lin +2
In deep reinforcement learning, building policies of high-quality is challenging when the feature space of states is small and the training data is limited. Despite the success of…