1 Sutton, R.S, "Reinforcement Learning : An Introduction" MIT press 1998
2 Christopher J. C. H. Watkins, "Q-learning" Springer Nature 8 (8): 279-292, 1992
3 Sutton, R.S, "Policy Gradient Methods for Reinforcement Learning with Function Approximation" 99 : 1057-1063, 1999
4 Fukushima, K, "Neocognitron : A Self-organizing Neural Network Model for a Mechanism of Pattern Recognition Unaffected by Shift in Position" 34 (34): 193-202, 1980
5 Mnih, V, "Human-level Control Through Deep Reinforcement Learning" 518 (518): 529-533, 2015
6 Bengio, Y, "Greedy Layer-wise Training of Deep Networks" 19 : 153-, 2007
7 Hsu, K, "Artificial Neural Network Modeling of the Rainfall Runoff Process" 31 (31): 2517-2530, 1995
1 Sutton, R.S, "Reinforcement Learning : An Introduction" MIT press 1998
2 Christopher J. C. H. Watkins, "Q-learning" Springer Nature 8 (8): 279-292, 1992
3 Sutton, R.S, "Policy Gradient Methods for Reinforcement Learning with Function Approximation" 99 : 1057-1063, 1999
4 Fukushima, K, "Neocognitron : A Self-organizing Neural Network Model for a Mechanism of Pattern Recognition Unaffected by Shift in Position" 34 (34): 193-202, 1980
5 Mnih, V, "Human-level Control Through Deep Reinforcement Learning" 518 (518): 529-533, 2015
6 Bengio, Y, "Greedy Layer-wise Training of Deep Networks" 19 : 153-, 2007
7 Hsu, K, "Artificial Neural Network Modeling of the Rainfall Runoff Process" 31 (31): 2517-2530, 1995