2 papers
cs.LG2021
EnTRPO: Trust Region Policy Optimization Method with Entropy Regularization
Sahar Roostaie, Mohammad Mehdi Ebadzadeh
Trust Region Policy Optimization (TRPO) is a popular and empirically successful policy search algorithm in reinforcement learning (RL). It iteratively solved the surrogate problem…
cs.LG2020
Adaptive Explicit Kernel Minkowski Weighted K-means
Amir Aradnia, Maryam Amir Haeri, Mohammad Mehdi Ebadzadeh
The K-means algorithm is among the most commonly used data clustering methods. However, the regular K-means can only be applied in the input space and it is applicable when cluster…