论文信息 - Neural Response Ranking for Social Conversation: A Data-Efficient Approach

Neural Response Ranking for Social Conversation: A Data-Efficient Approach

The overall objective of ‘social’ dialogue systems is to support engaging, entertaining, and lengthy conversations on a wide variety of topics, including social chit-chat. Apart from raw dialogue data, user-provided ratings are the most common signal used to train such systems to produce engaging responses. In this paper we show that social dialogue systems can be trained effectively from raw unannotated data. Using a dataset of real conversations collected in the 2017 Alexa Prize challenge, we developed a neural ranker for selecting ‘good’ system responses to user utterances, i.e. responses which are likely to lead to long and engaging conversations. We show that (1) our neural ranker consistently outperforms several strong baselines when trained to optimise for user ratings; (2) when trained on larger amounts of data and only using conversation length as the objective, the ranker performs better than the one trained using ratings – ultimately reaching a Precision@1 of 0.87. This advance will make data collection for social conversational agents simpler and less expensive in the future.

Oliver Lemon | Ondrej Dusek | Igor Shalyminov

[1] John Langford,et al. A reliable effective terascale linear learning system , 2011, J. Mach. Learn. Res..

[2] Zhou Yu,et al. TickTock: A Non-Goal-Oriented Multimodal Dialog System with Engagement Awareness , 2015, AAAI Spring Symposia.

[3] Gregory N. Hullender,et al. Learning to rank using gradient descent , 2005, ICML.

[4] Christopher D. Manning,et al. Incorporating Non-local Information into Information Extraction Systems by Gibbs Sampling , 2005, ACL.

[5] Denis G. Fedorenko,et al. Avoiding Echo-Responses in a Retrieval-Based Conversation System , 2017, ArXiv.

[6] Eric Gilbert,et al. VADER: A Parsimonious Rule-Based Model for Sentiment Analysis of Social Media Text , 2014, ICWSM.

[7] Yoshua Bengio,et al. Deep Sparse Rectifier Neural Networks , 2011, AISTATS.

[8] Jeff Johnson,et al. Billion-Scale Similarity Search with GPUs , 2017, IEEE Transactions on Big Data.

[9] Milica Gasic,et al. The Hidden Information State model: A practical framework for POMDP-based spoken dialogue management , 2010, Comput. Speech Lang..

[10] Oliver Lemon,et al. An Ensemble Model with Ranking for Social Dialogue , 2017, NIPS 2017.

[11] Yichao Lu,et al. A practical approach to dialogue response generation in closed domains , 2017, ArXiv.