62 citations · 381 across the 36 of their papers we have counts for
3 papers · 1 filter
Towards Efficient Detection and Optimal Response against Sophisticated Opponents
Tianpei Yang, Zhaopeng Meng, Jianye Hao +3
Multiagent algorithms often aim to accurately predict the behaviors of other agents and find a best response accordingly. Previous works usually assume an opponent uses a stationar…
Object-Oriented Dynamics Predictor
Guangxiang Zhu, Zhiao Huang, Chongjie Zhang
Generalization has been one of the major challenges for learning dynamics models in model-based reinforcement learning. However, previous work on action-conditioned dynamics predic…
Context-Aware Policy Reuse
Siyuan Li, Fangda Gu, Guangxiang Zhu +1
Transfer learning can greatly speed up reinforcement learning for a new task by leveraging policies of relevant tasks. Existing works of policy reuse either focus on only selecting…