1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2024★ 1 cited
Cross-Domain Policy Transfer by Representation Alignment via Multi-Domain Behavioral Cloning
Hayato Watahiki, Ryo Iwase, Ryosuke Unno +1
Transferring learned skills across diverse situations remains a fundamental challenge for autonomous agents, particularly when agents are not allowed to interact with an exact targ…
cs.LG2019
Reconnaissance and Planning algorithm for constrained MDP
Shin-ichi Maeda, Hayato Watahiki, Shintarou Okada +1
Practical reinforcement learning problems are often formulated as constrained Markov decision process (CMDP) problems, in which the agent has to maximize the expected return while…