3 papers
stat.ML2016
Policy Search with High-Dimensional Context Variables
Voot Tangkaratt, Herke van Hoof, Simone Parisi +3
Direct contextual policy search methods learn to improve policy parameters and simultaneously generalize these parameters to different context or task variables. However, learning…
cs.RO2014
Efficient Reuse of Previous Experiences to Improve Policies in Real Environment
Norikazu Sugimoto, Voot Tangkaratt, Thijs Wensveen +3
In this study, we show that a movement policy can be improved efficiently using the previous experiences of a real robot. Reinforcement Learning (RL) is becoming a popular approach…
cs.LG2014
Conditional Density Estimation with Dimensionality Reduction via Squared-Loss Conditional Entropy Minimization
Voot Tangkaratt, Ning Xie, Masashi Sugiyama
Regression aims at estimating the conditional mean of output given input. However, regression is not informative enough if the conditional density is multimodal, heteroscedastic, a…