Recent Machine Learning Papers

Advanced Search
2019-08-27

Deep Reinforcement Learning for Chatbots Using Clustered Actions and Human-Likeness Rewards

Heriberto Cuayáhuitl, Donghyeon Lee, Seonghan Ryu, Sungja Choi, Inchul Hwang, Jihie Kim

Training chatbots using the reinforcement learning paradigm is challenging due to high-dimensional states, infinite action spaces and the difficulty in specifying the reward function. We address such problems using clustered actions instead of infinite actions, and a simple but promising reward func...