most citedV-MPO: On-Policy Maximum a Posteriori Policy Optimization for Discrete and Continuous Control

39 citations · 158 across the 9 of their papers we have counts for

collaborators

9 papers

cs.LG202024 cited

A Distributional View on Multi-Objective Policy Optimization

Abbas Abdolmaleki, Sandy H. Huang, Leonard Hasenclever +7

Many real-world problems require trading off multiple competing objectives. However, these objectives are often in different units and/or scales, which can make it challenging for…

cs.LG202027 cited

Continuous-Discrete Reinforcement Learning for Hybrid Control in Robotics

Michael Neunert, Abbas Abdolmaleki, Markus Wulfmeier +7

Many real-world control problems involve both discrete decision variables - such as the choice of control modes, gear switching or digital outputs - as well as continuous decision…

cs.LG20194 cited

Quinoa: a Q-function You Infer Normalized Over Actions

Jonas Degrave, Abbas Abdolmaleki, Jost Tobias Springenberg +2

We present an algorithm for learning an approximate action-value soft Q-function in the relative entropy regularised reinforcement learning setting, for which an optimal improved p…

cs.RO201916 cited

Modelling Generalized Forces with Reinforcement Learning for Sim-to-Real Transfer

Rae Jeong, Jackie Kay, Francesco Romano +6

Learning robotic control policies in the real world gives rise to challenges in data efficiency, safety, and controlling the initial condition of the system. On the other hand, sim…

cs.RO20196 cited

Imagined Value Gradients: Model-Based Policy Optimization with Transferable Latent Dynamics Models

Arunkumar Byravan, Jost Tobias Springenberg, Abbas Abdolmaleki +6

Humans are masters at quickly learning many complex tasks, relying on an approximate understanding of the dynamics of their environments. In much the same way, we would like our le…

cs.LG20196 cited

Augmenting learning using symmetry in a biologically-inspired domain

Shruti Mishra, Abbas Abdolmaleki, Arthur Guez +2

Invariances to translation, rotation and other spatial transformations are a hallmark of the laws of motion, and have widespread use in the natural sciences to reduce the dimension…