Fetching the paper…
Reading the bibliography…
With the breakthrough of computational power and deep neural networks, many areas that we haven't explore with various techniques that was researched rigorously in past is feasible.
R. Bellman, Dynamic Programming, Princeton University Press, Princeton, NJ, 1957
1957
Earlier work this paper cites.
J.A. Nelder, R. Mead, A Simplex Method for Function Minimization, Computer Journal, Vol. 7, Issue 4, 1965
1965
Earlier work this paper cites.
R. Sutton, Learning to Predict By the Method of Temporal Differences, Machine Learing, vol.3, 1988
1988
Earlier work this paper cites.
P.J. Werbos, Back-Propagation Through Time: What it Does and How to Do it, IEEE Proceedings, vol.78, Oct, 1990
1990
Earlier work this paper cites.
R. Sutton, A. Barto, R. Williams, Reinforcement Learning is Direct Adaptive Optimal Control, IEEE Control Systems, April, 1992
1992
Earlier work this paper cites.
C.J. Watkins, P.Dayan, Technical note: Q-Learning, Machine Learning, vol.9, 1992
1992
Earlier work this paper cites.
W. Sharpe, The Sharpe Ratio, Journal of Portfolio Management, vol.21, 1994
1994
Earlier work this paper cites.
Y. Bengio, P. Simard, P. Frasconi, Learning Long-Term Dependencies with Gradient Descent is Difficult, IEEE Transactions On Neural Networks, March, 1994
1994
Earlier work this paper cites.
S. Boyd, L. El Ghaoui, E. Feron, V. Balakrishnan, Linear Matrix Inequilities in System and Control Theory, Society for Industrial and Applied Mathematics, 1994
1994
Earlier work this paper cites.
H.P. Schwefel: Evolution and Optimum Seeking: New York: Wiley & Sons 1995
1995
Earlier work this paper cites.
L. Kaelbling, M. Littman, A. Moore, Reinforcment Learning” A Survey, Journal of Artificial Intelligence Research, vol.4, 1996
1996
Earlier work this paper cites.
S. Hochreiter, J. Schmidhuber, Long Short-Term Memory, Neural Computation, 1997
1997
Cited alongside, same era.
J. Moody, L. Wu, Y. Liao, M. Saffell, Performance Functions and Reinforcement Learning for Trading Systems and Portfolios, Journal of Forecasting, vol. 17, 1998
1998
Cited alongside, same era.
F. Gers, N. Schraudolph, J. Schmidhuber, Learning to forget: Continual prediction with LSTM, Neural Computation, 2000
2000
Cited alongside, same era.
J. Moody, M. Saffell, Learning to Trade via Direct Reinforcement, IEEE Transactions on Neural Networks, Vol.12, July, 2001
2001
Cited alongside, same era.
H. Beyer, The Theory of Evolution Strategies, Springer-Verlag, New York, USA, 2001
2001
Cited alongside, same era.
J. Schmidhuber, D. Wiestra, M. Gagliolo, F. Gomoze, Training Recurrent Networks by Evolino, Neural Computation, 19, 2007
2007
Later among the works it cites.
D. Wierstra, T. Schaul, J. Peters,J. Schmidhuber, Natural evolution strategies, In Proceedings of the Congress on Evolutionary Computation, IEEE Press, 2008
2008
Later among the works it cites.
R. Sutton, A. Barto, Reinforcement Learning: An Introduction, 2nd ed. Cambridge, Massachusetts, The MIT Press, 2012
2012
Later among the works it cites.
Y. Kao, B. Van Roy, Directed Principal Component Analysis, Operations Research, July, 2014
2014
Later among the works it cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, R. Salakhutdinov, Dropout: A Simple Way to Prevent Neural Networks from Overfitting, Journal of Machine Learning Research, 15, 2014
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2001
Cited alongside, same era.
N. Hansen, A. Ostermeier, Completely Derandomized Self-Adaptation in Evolution Strategies, In Evolutionary Computation, 9, 2001
2001
Cited alongside, same era.
C. Gold, FX Trading via Recurrent Reinforcement Learning, Computation and Neural Systems, 2003
2003
Cited alongside, same era.
R. Y. Rubinstein, D. P. Kroese, The Cross-Entropy Method: A Unified Approach to Combinatorial Optimization, Monte-Carlo Simulation and Machine Learning (Information Science and Statistics), Springer, 2004
2004
Cited alongside, same era.
D. Lu, Portfolio Optimization Using Linear Matrix Inequilities, IIT Working Paper, 2005
2005
Cited alongside, same era.
C.J. Price, I.D. Coope, D. Byatt, A Convergent Variant of the Nelder-Maead Algorithm, J.Optim. Theory Appl. 113, No.1
Cited in the paper.
Y. Deng, F. Bao, Y. Kong, Z. Ren, Q. Dai, Deep Direct Reinforcement Learning for Financial Signal Representation and Trading, IEEE Transactions on Neural Networks and Learning Systems, April, 2015
2015
Later among the works it cites.
W. Zaremba, I. Sutskever, O. Vinyals, Recurrent Neural Network Regularization, ICLR, 2015
2015
Later among the works it cites.
I. Goodfellow, Y. Bengio, A. Courville, Deep Learning, Book Draft for The MIT Press, 2015
2015
Later among the works it cites.
I. Danihelka, G. Wayne, B. Uria, Nal Kalchbrenner, Alex Graves, Associative Long Short-Term Memory, arXiv, 2015
2015
Later among the works it cites.
W. Zaremba, I. Sutskever, O. Vinyals, Recurrent Neural Network Regularization, arXiv, Feb, 2015
2015
Later among the works it cites.