Alle Publikationen

  • Erweiterte Suche öffnen

Treffer: 1
  • <<
  • 1
  • 2015

  • Ferreira, Emmanuel; Lefèvre, Fabrice (2015): Reinforcement-learning based dialogue system for human–robot interactions with socially-inspired rewards. In: Computer Speech & Language 34 (1), S. 256-274. DOI: 10.1016/j.csl.2015.03.007

    DOI: https://doi.org/10.1016/j.csl.2015.03.007 

    Abstract: This paper investigates some conditions under which polarized user appraisals gathered throughout the course of a vocal interaction between a machine and a human can be integrated in a reinforcement learning-based dialogue manager. More specifically, we discuss how this information can be cast into socially-inspired rewards for speeding up the policy optimisation for both efficient task completion and user adaptation in an online learning setting. For this purpose a potential-based reward shaping method is combined with a sample efficient reinforcement learning algorithm to offer a principled framework to cope with these potentially noisy interim rewards. The proposed scheme will greatly facilitate the system's development by allowing the designer to teach his system through explicit positive/negative feedbacks given as hints about task progress, in the early stage of training. At a later stage, the approach will be used as a way to ease the adaptation of the dialogue policy to specific user profiles. Experiments carried out using a state-of-the-art goal-oriented dialogue management framework, the Hidden Information State (HIS), support our claims in two configurations: firstly, with a user simulator in the tourist information domain (and thus simulated appraisals), and secondly, in the context of man–robot dialogue with real user trials.

  • <<
  • 1