Generative Deep Neural Networks for Dialogue: A Short Review
Iulian Vlad Serban Affiliation: Department of Computer Science Affiliation: and Operations Research, Affiliation: University of Montreal Ryan Lowe Affiliation: School of Computer Science, Affiliation: McGill University Laurent Charlin Affiliation: School of Computer Science, Affiliation: McGill University Joelle Pineau Affiliation: School of Computer Science, Affiliation: McGill University
Abstract
Researchers have recently started investigating deep neural networks for dialogue applications. In particular, generative sequence-to-sequence (Seq2Seq) models have shown promising results for unstructured tasks, such as word-level dialogue response generation. The hope is that such models will be able to leverage massive amounts of data to learn meaningful natural language representations and response generation strategies, while requiring a minimum amount of domain knowledge and hand-crafting. An important challenge is to develop models that can effectively incorporate dialogue context and generate meaningful and diverse responses. In support of this goal, we review recently proposed models based on generative encoder-decoder neural network architectures, and show that these models have better ability to incorporate long-term dialogue history, to model uncertainty and ambiguity in dialogue, and to generate responses with high-level compositional structure.
原文 arXiv:1611.06216;中英对照 + 大白话阅读 https://aha.fim.ai/paper/1611.06216v1