2 papers
cs.HC2018
DQN-TAMER: Human-in-the-Loop Reinforcement Learning with Intractable Feedback
Riku Arakawa, Sosuke Kobayashi, Yuya Unno +2
Exploration has been one of the greatest challenges in reinforcement learning (RL), which is a large obstacle in the application of RL to robotics. Even with state-of-the-art RL al…
cs.CL2018
Addressee and Response Selection for Multilingual Conversation
Motoki Sato, Hiroki Ouch, Yuta Tsuboi
Developing conversational systems that can converse in many languages is an interesting challenge for natural language processing. In this paper, we introduce multilingual addresse…