论文信息 - Playing Text-Adventure Games with Graph-Based Deep Reinforcement Learning

Playing Text-Adventure Games with Graph-Based Deep Reinforcement Learning

Text-based adventure games provide a platform on which to explore reinforcement learning in the context of a combinatorial action space, such as natural language. We present a deep reinforcement learning architecture that represents the game state as a knowledge graph which is learned during exploration. This graph is used to prune the action space, enabling more efficient exploration. The question of which action to take can be reduced to a question-answering task, a form of transfer learning that pre-trains certain parts of our architecture. In experiments using the TextWorld framework, we show that our proposed technique can learn a control policy faster than baseline alternatives. We have also open-sourced our code at https://github.com/rajammanabrolu/KG-DQN.

Mark O. Riedl | Prithviraj Ammanabrolu | Prithviraj Ammanabrolu

[1] Shie Mannor,et al. Learning How Not to Act in Text-based Games , 2018, ICLR.

[2] Pietro Liò,et al. Graph Attention Networks , 2017, ICLR.

[3] Richard S. Sutton,et al. Reinforcement Learning: An Introduction , 1998, IEEE Trans. Neural Networks.

[4] David Wingate,et al. What Can You Do with a Rock? Affordance Extraction via Word Embeddings , 2017, IJCAI.

[5] Christopher D. Manning,et al. Leveraging Linguistic Structure For Open Domain Information Extraction , 2015, ACL.

[6] Andrew W. Moore,et al. Prioritized sweeping: Reinforcement learning with less data and less time , 2004, Machine Learning.

[7] Jason Weston,et al. Reading Wikipedia to Answer Open-Domain Questions , 2017, ACL.

[8] Minlie Huang,et al. Story Ending Generation with Incremental Encoding and Commonsense Knowledge , 2018, AAAI.

[9] Quoc V. Le,et al. Sequence to Sequence Learning with Neural Networks , 2014, NIPS.

[10] Marc-Alexandre Côté,et al. Towards Solving Text-based Games by Producing Adaptive Action Spaces , 2018, ArXiv.

[11] Richard Socher,et al. The Natural Language Decathlon: Multitask Learning as Question Answering , 2018, ArXiv.