Search CORE

272,956 research outputs found

Deep reinforcement learning for multi-domain dialogue systems

Author: Carse Jacob
Cuayahuitl Heriberto
Williamson Ashley
Yu Seunghak
Publication venue: 'Center for Open Science'
Publication date: 26/11/2016
Field of study

Standard deep reinforcement learning methods such as Deep Q-Networks (DQN) for multiple tasks (domains) face scalability problems. We propose a method for multi-domain dialogue policy learning---termed NDQN, and apply it to an information-seeking spoken dialogue system in the domains of restaurants and hotels. Experimental results comparing DQN (baseline) versus NDQN (proposed) using simulations report that our proposed method exhibits better scalability and is promising for optimising the behaviour of multi-domain dialogue systems

University of Lincoln Institutional Repository

arXiv.org e-Print Archive

End-to-end optimization of goal-driven and visually grounded dialogue systems

Author: Courville Aaron
de Vries Harm
Mary Jeremie
Pietquin Olivier
Piot Bilal
Strub Florian
Publication venue
Publication date: 15/03/2017
Field of study

End-to-end design of dialogue systems has recently become a popular research topic thanks to powerful tools such as encoder-decoder architectures for sequence-to-sequence learning. Yet, most current approaches cast human-machine dialogue management as a supervised learning problem, aiming at predicting the next utterance of a participant given the full history of the dialogue. This vision is too simplistic to render the intrinsic planning problem inherent to dialogue as well as its grounded nature, making the context of a dialogue larger than the sole history. This is why only chit-chat and question answering tasks have been addressed so far using end-to-end architectures. In this paper, we introduce a Deep Reinforcement Learning method to optimize visually grounded task-oriented dialogues, based on the policy gradient algorithm. This approach is tested on a dataset of 120k dialogues collected through Mechanical Turk and provides encouraging results at solving both the problem of generating natural dialogues and the task of discovering a specific object in a complex picture

arXiv.org e-Print Archive

Crossref

INRIA a CCSD electronic archive server

HAL Descartes

Hal-Diderot