Learning to Start for Sequence to Sequence Based Response Generation

Qingfu Zhu,Weinan Zhang,Ting Liu
DOI: https://doi.org/10.1007/978-3-030-01012-6_22
2018-01-01
Abstract:Response Generation which is a crucial component of a dialogue system can be modeled using the Sequence to Sequence (Seq2Seq) architecture. However, this kind of method suffers from vague responses of little meaningful content. One possible reason for generating vague responses is the different distribution of the first word between the generated responses and human responses. In fact, the Seq2Seq based method tends to generate high-frequency words in the beginning, which influences the following prediction resulting in vague responses. In this paper, we proposed a novel approach, namely learning to start (LTS), to learn how to generate the first word in the sequence to sequence architecture for response generation. Experimental results show that the proposed LTS model can enhance the performance of the start-of-the-art Seq2Seq model as well as other Seq2Seq models for response generation of short text conversation.
What problem does this paper attempt to address?