6 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.RO2023★ 6 cited
Affordance-Driven Next-Best-View Planning for Robotic Grasping
Xuechao Zhang, Dong Wang, Sun Han +7
Grasping occluded objects in cluttered environments is an essential component in complex robotic manipulation tasks. In this paper, we introduce an AffordanCE-driven Next-Best-View…
cs.LG2023
Revisiting Estimation Bias in Policy Gradients for Deep Reinforcement Learning
Haoxuan Pan, Deheng Ye, Xiaoming Duan +4
We revisit the estimation bias in policy gradients for the discounted episodic Markov decision process (MDP) from Deep Reinforcement Learning (DRL) perspective. The objective is fo…
math.OC2021★ 1 cited
On the Detection of Markov Decision Processes
Xiaoming Duan, Yagiz Savas, Rui Yan +2
We study the detection problem for a finite set of Markov decision processes (MDPs) where the MDPs have the same state and action spaces but possibly different probabilistic transi…