4 citations · 8 across the 4 of their papers we have counts for
4 papers
Sample Efficient Grasp Learning Using Equivariant Models
Xupeng Zhu, Dian Wang, Ondrej Biza +3
In planar grasp detection, the goal is to learn a function from an image of a scene onto a set of feasible grasp poses in . In this paper, we recognize that the opt…
Equivariant Learning in Spatial Action Spaces
Dian Wang, Robin Walters, Xupeng Zhu +1
Recently, a variety of new equivariant neural network model architectures have been proposed that generalize better over rotational and reflectional symmetries than standard models…
Lipschitz Bandit Optimization with Improved Efficiency
Xu Zhu
We consider the Lipschitz bandit optimization problem with an emphasis on practical efficiency. Although there is rich literature on regret analysis of this type of problem, e.g.,…
Stochastic Lipschitz Q-Learning
Xu Zhu
In an episodic Markov Decision Process (MDP) problem, an online algorithm chooses from a set of actions in a sequence of trials, where is the episode length, in order to ma…