2 papers
cs.RO2025
Demonstrating Multi-Suction Item Picking at Scale via Multi-Modal Learning of Pick Success
Che Wang, Jeroen van Baar, Chaitanya Mitash +6
This work demonstrates how autonomously learning aspects of robotic operation from sparsely-labeled, real-world data of deployed, engineered solutions at industrial scale can provi…
cs.LG2025
On the Convergence of Monte Carlo UCB for Random-Length Episodic MDPs
Zixuan Dong, Che Wang, Keith Ross
In reinforcement learning, Monte Carlo algorithms update the Q function by averaging the episodic returns. In the Monte Carlo UCB (MC-UCB) algorithm, the action taken in each state…