most citedTime manipulation technique for speeding up reinforcement learning in simulations

7 citations · 12 across the 6 of their papers we have counts for

collaborators

6 papers

cs.RO20094 cited

Intent expression using eye robot for mascot robot system

Yoichi Yamazaki, Fangyan Dong, Yuta Masuda +5

An intent expression system using eye robots is proposed for a mascot robot system from a viewpoint of humatronics. The eye robot aims at providing a basic interface method for an…

cs.RO2009

Fuzzy inference based mentality estimation for eye robot agent

Yoichi Yamazaki, Fangyan Dong, Yuta Masuda +5

Household robots need to communicate with human beings in a friendly fashion. To achieve better understanding of displayed information, an importance and a certainty of the informa…

cs.AI2009

Eligibility Propagation to Speed up Time Hopping for Reinforcement Learning

Petar Kormushev, Kohei Nomoto, Fangyan Dong +1

A mechanism called Eligibility Propagation is proposed to speed up the Time Hopping technique used for faster Reinforcement Learning in simulations. Eligibility Propagation provide…

cs.IR20091 cited

Visual approach for data mining on medical information databases using Fastmap algorithm

Petar Kormushev

The rapid development of tools for acquisition and storage of information has lead to the formation of enormous medical databases. The large quantity of data definitely surpasses t…

cs.AI2009

Design, development and implementation of a tool for construction of declarative functional descriptions of semantic web services based on WSMO methodology

Petar Kormushev

Semantic web services (SWS) are self-contained, self-describing, semantically marked-up software resources that can be published, discovered, composed and executed across the Web i…

cs.AI20097 cited

Time manipulation technique for speeding up reinforcement learning in simulations

Petar Kormushev, Kohei Nomoto, Fangyan Dong +1

A technique for speeding up reinforcement learning algorithms by using time manipulation is proposed. It is applicable to failure-avoidance control problems running in a computer s…