First Results from Using Temporal Difference Learning in Shogi
暂无分享,去创建一个
[1] Gerald Tesauro,et al. TD-Gammon, a Self-Teaching Backgammon Program, Achieves Master-Level Play , 1994, Neural Computation.
[2] Richard E. Korf,et al. A Unified Theory of Heuristic Evaluation Functions and its Application to Learning , 1986, AAAI.
[3] Gerald Tesauro,et al. Practical Issues in Temporal Difference Learning , 1992, Mach. Learn..
[4] Andrew Tridgell,et al. KnightCap: A Chess Programm That Learns by Combining TD(lambda) with Game-Tree Search , 1998, ICML.
[5] Donald F. Beal,et al. Learning Piece Values Using Temporal Differences , 1997, J. Int. Comput. Games Assoc..
[6] Robert Levinson,et al. Adaptive Pattern-Oriented Chess , 1991, AAAI Conference on Artificial Intelligence.
[7] Christian Donninger,et al. Null Move and Deep Search , 1993, J. Int. Comput. Games Assoc..
[8] J. Fairbairn. Shogi for beginners , 1984 .
[9] Tony Marsland,et al. COMPUTER CHESS AND SEARCH , 1992 .
[10] Hiroyuki Iida,et al. Natural Developments in Game Research , 1996, J. Int. Comput. Games Assoc..