Fetching the paper…
Reading the bibliography…
This paper addresses the general problem of reinforcement learning (RL) in partially observable environments.
In F. Hasenöhrl, editor, Wissenschaftliche Abhandlungen (collection of Boltzmann’s articles in scientific journals)
L. Boltzmann · 1909
Earlier work this paper cites.
Über formal unentscheidbare Sätze der Principia Mathematica und verwandter Systeme I
K. Gödel · 1931
Earlier work this paper cites.
An unsolvable problem of elementary number theory
A. Church · 1936
Earlier work this paper cites.
Finite combinatory processes-formulation 1
E. L. Post · 1936
Earlier work this paper cites.
On computable numbers, with an application to the Entscheidungsproblem
A. M. Turing · 1936
Earlier work this paper cites.
A mathematical theory of communication (parts I and II)
C. E. Shannon · 1948
Earlier work this paper cites.
On information and sufficiency
S. Kullback and R. A. Leibler · 1951
Earlier work this paper cites.
A method for construction of minimum-redundancy codes
D. A. Huffman · 1952
Earlier work this paper cites.
Dynamic Programming
R. Bellman · 1957
Earlier work this paper cites.
Gradient theory of optimal flight paths
H. J. Kelley · 1960
Earlier work this paper cites.
A gradient method for optimizing multi-stage allocation processes
A. E. Bryson · 1961
Earlier work this paper cites.
The numerical solution of variational problems
S. E. Dreyfus · 1962
Earlier work this paper cites.
A formal theory of inductive inference. Part I
R. J. Solomonoff · 1964
Earlier work this paper cites.
Cybernetic Predicting Devices
A. G. Ivakhnenko and V. G. Lapa · 1965
Earlier work this paper cites.
Three approaches to the quantitative definition of information
A. N. Kolmogorov · 1965
Earlier work this paper cites.
On the length of programs for computing finite binary sequences
G. J. Chaitin · 1966
Earlier work this paper cites.
Artificial Intelligence through Simulated Evolution
L. Fogel, A. Owens, and M. Walsh · 1966
Earlier work this paper cites.
The group method of data handling – a rival of the method of stochastic approximation
A. G. Ivakhnenko · 1968
Earlier work this paper cites.
An information theoretic measure for classification
C. S. Wallace and D. M. Boulton · 1968
Earlier work this paper cites.
Statistical predictor identification
H. Akaike · 1970
Earlier work this paper cites.
The representation of the cumulative rounding error of an algorithm as a Taylor expansion of the local rounding errors
S. Linnainmaa · 1970
Earlier work this paper cites.
Polynomial theory of complex systems
A. G. Ivakhnenko · 1971
Earlier work this paper cites.
On the notion of a random sequence
L. A. Levin · 1973
Earlier work this paper cites.
Universal sequential search problems
L. A. Levin · 1973
Earlier work this paper cites.
Evolutionsstrategie - Optimierung technischer Systeme nach Prinzipien der biologischen Evolution. Dissertation, 1971
I. Rechenberg · 1973
Earlier work this paper cites.
A new look at the statistical model identification
H. Akaike · 1974
Earlier work this paper cites.
Adaptation in Natural and Artificial Systems
J. H. Holland · 1975
Earlier work this paper cites.
Numerische Optimierung von Computer-Modellen. Dissertation, 1974
H. P. Schwefel · 1977
Earlier work this paper cites.
Complexity-based induction systems
R. J. Solomonoff · 1978
Earlier work this paper cites.
Neural network model for a mechanism of pattern recognition unaffected by shift in position - Neocognitron
K. Fukushima · 1979
Earlier work this paper cites.
Compiling Fast Partial Derivatives of Functions Given by Algorithms
B. Speelpenning · 1980
Earlier work this paper cites.
Applications of advances in nonlinear sensitivity analysis
P. J. Werbos · 1981
Earlier work this paper cites.
Neuronlike adaptive elements that can solve difficult learning control problems
A. G. Barto, R. S. Sutton, and C. W. Anderson · 1983
Earlier work this paper cites.
Learning while searching in constraint-satisfaction problems
R. Dechter · 1986
Earlier work this paper cites.
Stochastic complexity and modeling
J. Rissanen · 1986
Earlier work this paper cites.
Learning internal representations by error propagation
D. E. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
Reinforcement-learning in connectionist networks: A mathematical analysis
R. J. Williams · 1986
Earlier work this paper cites.
Modular learning in neural networks
D. H. Ballard · 1987
Earlier work this paper cites.
A dual back-propagation scheme for scalar reinforcement learning
P. W. Munro · 1987
Earlier work this paper cites.
The utility driven dynamic error propagation network
A. J. Robinson and F. Fallside · 1987
Earlier work this paper cites.
Building and understanding adaptive systems: A statistical/numerical approach to factory automation and brain research
P. J. Werbos · 1987
Earlier work this paper cites.
Connectionist expert systems
S. I. Gallant · 1988
Earlier work this paper cites.
A network of neuron-like units that learns to perceive by generation as well as reweighting of its links
V. Honavar and L. M. Uhr · 1988
Earlier work this paper cites.
Supervised learning and systems with excess degrees of freedom
M. I. Jordan · 1988
Earlier work this paper cites.
Implications of recursive distributed representations
J. B. Pollack · 1988
Earlier work this paper cites.
Generalization of backpropagation with application to a recurrent gas market model
P. J. Werbos · 1988
Earlier work this paper cites.
Toward a theory of reinforcement-learning connectionist systems
R. J. Williams · 1988
Earlier work this paper cites.
Dynamic node creation in backpropagation neural networks
T. Ash · 1989
Earlier work this paper cites.
Finding minimum entropy codes
H. B. Barlow, T. P. Kaushal, and G. J. Mitchison · 1989
Earlier work this paper cites.
Genetic Algorithms in Search, Optimization and Machine Learning
D. E. Goldberg · 1989
Earlier work this paper cites.
Comparing biases for minimal network construction with back-propagation
S. J. Hanson and L. Y. Pratt · 1989
Earlier work this paper cites.
Back-propagation applied to handwritten zip code recognition
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel · 1989
Earlier work this paper cites.
Designing neural networks using genetic algorithms
G. Miller, P. Todd, and S. Hedge · 1989
Earlier work this paper cites.
Fast learning in multi-resolution hierarchies
J. E. Moody · 1989
Earlier work this paper cites.
Skeletonization: A technique for trimming the fat from a network via relevance assessment
M. C. Mozer and P. Smolensky · 1989
Earlier work this paper cites.
The truck backer-upper: An example of self learning in neural networks
N. Nguyen and B. Widrow · 1989
Earlier work this paper cites.
Dynamic reinforcement driven error propagation networks with application to game playing
T. Robinson and F. Fallside · 1989
Earlier work this paper cites.
A local learning algorithm for dynamic feedforward and recurrent networks
J. Schmidhuber · 1989
Earlier work this paper cites.
Learning from Delayed Rewards
C. J. C. H. Watkins · 1989
Earlier work this paper cites.
Backpropagation and neurocontrol: A review and prospectus
P. J. Werbos · 1989
Earlier work this paper cites.
Neural networks for control and system identification
P. J. Werbos · 1989
Earlier work this paper cites.
Learning in artificial neural networks: A statistical perspective
H. White · 1989
Earlier work this paper cites.
Supervised learning with a distal teacher
M. I. Jordan and D. E. Rumelhart · 1990
Earlier work this paper cites.
Designing neural networks using genetic algorithms with graph generation system
H. Kitano · 1990
Earlier work this paper cites.
Optimal brain damage
Y. LeCun, J. S. Denker, and S. A. Solla · 1990
Earlier work this paper cites.
Identification and control of dynamical systems using neural networks
K. S. Narendra and K. Parthasarathy · 1990
Earlier work this paper cites.
The strength of weak learnability
R. E. Schapire · 1990
Earlier work this paper cites.
Learning algorithms for networks with internal and external feedback
J. Schmidhuber · 1990
Earlier work this paper cites.
An on-line algorithm for dynamic reinforcement learning and planning in reactive environments
J. Schmidhuber · 1990
Earlier work this paper cites.
Integrated architectures for learning, planning and reacting based on dynamic programming
R. S. Sutton · 1990
Earlier work this paper cites.
The recurrent cascade-correlation learning algorithm
S. E. Fahlman · 1991
Earlier work this paper cites.
Introduction to the Theory of Neural Computation
J. Hertz, A. Krogh, and R. Palmer · 1991
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen Netzen. Diploma thesis, Institut für Informatik, Lehrstuhl Prof. Brauer, Technische Universität München, 1991
S. Hochreiter · 1991
Earlier work this paper cites.
Delayed reinforcement learning with multiple time scale hierarchical backpropagated adaptive critics
J. Jameson · 1991
Earlier work this paper cites.
Programming robots using reinforcement learning and teaching
L.-J. Lin · 1991
Earlier work this paper cites.
Incremental development of complex behaviors through automatic construction of sensory-motor hierarchies
M. B. Ring · 1991
Earlier work this paper cites.
Curious model-building control systems
J. Schmidhuber · 1991
Earlier work this paper cites.
Learning to generate sub-goals for action sequences
J. Schmidhuber · 1991
Earlier work this paper cites.
A possibility for implementing curiosity and boredom in model-building neural controllers
J. Schmidhuber · 1991
Earlier work this paper cites.
Reinforcement learning in Markovian and non-Markovian environments
J. Schmidhuber · 1991
Earlier work this paper cites.
Learning complex, extended sequences using the principle of history compression
J. Schmidhuber · 1991
Earlier work this paper cites.
Learning to generate artificial fovea trajectories for target detection
J. Schmidhuber and R. Huber · 1991
Earlier work this paper cites.
Turing computability with neural nets
H. T. Siegelmann and E. D. Sontag · 1991
Earlier work this paper cites.
Generalization by weight-elimination with application to forecasting
A. S. Weigend, D. E. Rumelhart, and B. A. Huberman · 1991
Earlier work this paper cites.
Evolving neural network controllers for unstable systems
A. P. Wieland · 1991
Earlier work this paper cites.
Application of time-bounded Kolmogorov complexity in complexity theory
A. Allender · 1992
Earlier work this paper cites.
Learning context-free grammars: Capabilities and limitations of a neural network with an external stack memory
S. Das, C. Giles, and G. Sun · 1992
Earlier work this paper cites.
Structural risk minimization for character recognition
I. Guyon, V. Vapnik, B. Boser, L. Bottou, and S. A. Solla · 1992
Earlier work this paper cites.
A simple weight decay can improve generalization
A. Krogh and J. A. Hertz · 1992
Earlier work this paper cites.
Memory approaches to reinforcement learning in non-markovian domains
L.-J. Lin and T. M. Mitchell · 1992
Earlier work this paper cites.
A practical Bayesian framework for backprop networks
D. J. C. MacKay · 1992
Earlier work this paper cites.
The effective number of parameters: An analysis of generalization and regularization in nonlinear learning systems
J. E. Moody · 1992
Earlier work this paper cites.
Learning to control fast-weight memories: An alternative to recurrent nets
J. Schmidhuber · 1992
Earlier work this paper cites.
Principles of risk minimization for learning theory
V. Vapnik · 1992
Earlier work this paper cites.
Kolmogorov complexity and computational complexity
O. Watanabe · 1992
Earlier work this paper cites.
Q-learning
C. J. C. H. Watkins and P. Dayan · 1992
Earlier work this paper cites.
Cresceptron: a self-organizing neural network which grows adaptively
J. Weng, N. Ahuja, and T. S. Huang · 1992
Earlier work this paper cites.
Neural networks, system identification, and control in the chemical industries
P. J. Werbos · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Statistical theory of learning curves under entropic loss criterion
S. Amari and N. Murata · 1993
Earlier work this paper cites.
Evolving recurrent dynamical networks for robot control
D. T. Cliff, P. Husbands, and I. Harvey · 1993
Earlier work this paper cites.
Neural networks for optimization and signal processing
A. Cochocki and R. Unbehauen · 1993
Earlier work this paper cites.
Feudal reinforcement learning
P. Dayan and G. Hinton · 1993
Earlier work this paper cites.
Neural network control for a closed-loop system using feedback-error-learning
H. Gomi and M. Kawato · 1993
Earlier work this paper cites.
Second order derivatives for network pruning: Optimal brain surgeon
B. Hassibi and D. G. Stork · 1993
Earlier work this paper cites.
Keeping neural networks simple
G. E. Hinton and D. van Camp · 1993
Earlier work this paper cites.
Generative learning structures and processes for generalized connectionist networks
V. Honavar and L. Uhr · 1993
Earlier work this paper cites.
Reinforcement Learning for Robots Using Neural Networks
L. Lin · 1993
Earlier work this paper cites.
Prioritized sweeping: Reinforcement learning with less data and less time
A. Moore and C. G. Atkeson · 1993
Cited alongside, same era.
A connectionist symbol manipulator that discovers the structure of context-free languages
M. C. Mozer and S. Das · 1993
Cited alongside, same era.
Learning sequential tasks by incrementally adding higher orders
M. B. Ring · 1993
Cited alongside, same era.
Netzwerkarchitekturen, Zielfunktionen und Kettenregel. (Network architectures, objective functions, and chain rule.)
J. Schmidhuber · 1993
Cited alongside, same era.
On decreasing the ratio between learning complexity and number of time-varying variables in fully recurrent nets
J. Schmidhuber · 1993
Cited alongside, same era.
A self-referential weight matrix
J. Schmidhuber · 1993
Hierarchical reinforcement learning based on subgoal discovery and subpolicy specialization
B. Bakker and J. Schmidhuber · 2004
Later among the works it cites.
Intrinsically motivated learning of hierarchical collections of skills
A. G. Barto, S. Singh, and N. Chentanez · 2004
Later among the works it cites.
Harnessing nonlinearity: Predicting chaotic systems and saving energy in wireless communication
H. Jaeger · 2004
Later among the works it cites.
A hybrid of genetic algorithm and particle swarm optimization for recurrent network design
C.-F. Juang · 2004
Later among the works it cites.
Policy gradient reinforcement learning for fast quadrupedal locomotion
N. Kohl and P. Stone · 2004
Later among the works it cites.
GPU implementation of neural networks
K.-S. Oh and K. Jung · 2004
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A reinforcement learning method for maximizing undiscounted rewards
A. Schwartz · 1993
Cited alongside, same era.
Learning via task decomposition
J. Tenenberg, J. Karlsson, and S. Whitehead · 1993
Cited alongside, same era.
A review of evolutionary artificial neural networks
X. Yao · 1993
Cited alongside, same era.
A constructive algorithm that converges for real-valued input patterns
N. Burgess · 1994
Cited alongside, same era.
A growing neural gas network learns topologies
B. Fritzke · 1994
Cited alongside, same era.
On the Theory of Generalization and Self-Structuring in Linearly Weighted Connectionist Networks
S. B. Holden · 1994
Cited alongside, same era.
Later among the works it cites.
Optimal ordered problem solver
J. Schmidhuber · 2004
Later among the works it cites.
Framewise phoneme classification with bidirectional LSTM and other neural network architectures
A. Graves and J. Schmidhuber · 2005
Later among the works it cites.
Advances in minimum description length: Theory and applications
P. D. Grünwald, I. J. Myung, and M. A. Pitt · 2005
Later among the works it cites.
Universal Artificial Intelligence: Sequential Decisions based on Algorithmic Probability
M. Hutter · 2005
Later among the works it cites.
Neural fitted Q iteration—first experiences with a data efficient neural reinforcement learning method
M. Riedmiller · 2005
Later among the works it cites.
Intrinsically motivated reinforcement learning
S. Singh, A. G. Barto, and N. Chentanez · 2005
Later among the works it cites.
Evolving keepaway soccer players through task decomposition
S. Whiteson, N. Kohl, R. Miikkulainen, and P. Stone · 2005
Later among the works it cites.
Pattern Recognition and Machine Learning
C. M. Bishop · 2006
Later among the works it cites.
High performance convolutional neural networks for document processing
K. Chellapilla, S. Puri, and P. Simard · 2006
Later among the works it cites.
Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural nets
A. Graves, S. Fernandez, F. J. Gomez, and J. Schmidhuber · 2006
Later among the works it cites.
Reducing the dimensionality of data with neural networks
G. Hinton and R. Salakhutdinov · 2006
Later among the works it cites.
Learning long term dependencies with recurrent neural networks
A. M. Schäfer, S. Udluft, and H.-G. Zimmermann · 2006
Later among the works it cites.
Developmental robotics, optimal artificial curiosity, creativity, music, and the fine arts
J. Schmidhuber · 2006
Later among the works it cites.
Closed-loop learning of visual control policies
S. R. Jodogne and J. H. Piater · 2007
Later among the works it cites.
Unsupervised learning of invariant feature hierarchies with applications to object recognition
M. A. Ranzato, F. Huang, Y. Boureau, and Y. LeCun · 2007
Later among the works it cites.
Training recurrent networks by EVOLINO
J. Schmidhuber, D. Wierstra, M. Gagliolo, and F. J. Gomez · 2007
Later among the works it cites.
Accelerated neural evolution through cooperatively coevolved synapses
F. J. Gomez, J. Schmidhuber, and R. Miikkulainen · 2008
Later among the works it cites.
A system for robotic heart surgery that learns to tie knots using recurrent neural networks
H. Mayer, F. Gomez, D. Wierstra, I. Nagy, A. Knoll, and J. Schmidhuber · 2008
Later among the works it cites.
Natural actor-critic
J. Peters and S. Schaal · 2008
Later among the works it cites.
Reinforcement learning of motor skills with policy gradients
J. Peters and S. Schaal · 2008
Later among the works it cites.
State-Dependent Exploration for policy gradient methods
T. Rückstieß, M. Felder, and J. Schmidhuber · 2008
Later among the works it cites.
Skill characterization based on betweenness
Ö. Simsek and A. G. Barto · 2008
Later among the works it cites.
A convergent O(n) algorithm for off-policy temporal-difference learning with linear function approximation
R. S. Sutton, C. Szepesvári, and H. R. Maei · 2008
Later among the works it cites.
Natural evolution strategies
D. Wierstra, T. Schaul, J. Peters, and J. Schmidhuber · 2008
Later among the works it cites.
A novel connectionist system for improved unconstrained handwriting recognition
A. Graves, M. Liwicki, S. Fernandez, R. Bertolami, H. Bunke, and J. Schmidhuber · 2009
Later among the works it cites.
Offline handwriting recognition with multidimensional recurrent neural networks
A. Graves and J. Schmidhuber · 2009
Later among the works it cites.
The Intelligent Movement Machine: An Ethological Perspective on the Primate Motor System
M. Graziano · 2009
Later among the works it cites.
Neuroevolution strategies for episodic reinforcement learning
V. Heidrich-Meisner and C. Igel · 2009
Later among the works it cites.
Large-scale deep unsupervised learning using graphics processors
R. Raina, A. Madhavan, and A. Ng · 2009
Later among the works it cites.
Simple algorithmic theory of subjective beauty, novelty, surprise, interestingness, attention, curiosity, creativity, art, science, music, jokes
J. Schmidhuber · 2009
Later among the works it cites.
A hypercube-based encoding for evolving large-scale neural networks
K. O. Stanley, D. B. D’Ambrosio, and J. Gauci · 2009
Later among the works it cites.
Efficient natural evolution strategies
Y. Sun, D. Wierstra, T. Schaul, and J. Schmidhuber · 2009
Later among the works it cites.
Hierarchical controller learning in a first-person shooter
N. van Hoorn, J. Togelius, and J. Schmidhuber · 2009
Later among the works it cites.
Deep big simple neural nets for handwritten digit recogntion
D. C. Ciresan, U. Meier, L. M. Gambardella, and J. Schmidhuber · 2010
Later among the works it cites.
Stable adaptive neural network control
S. Ge, C. C. Hang, T. H. Lee, and T. Zhang · 2010
Later among the works it cites.
Exponential natural evolution strategies
T. Glasmachers, T. Schaul, Y. Sun, D. Wierstra, and J. Schmidhuber · 2010
Later among the works it cites.
Multi-Dimensional Deep Memory Atari-Go Players for Parameter Exploring Policy Gradients
M. Grüttner, F. Sehnke, T. Schaul, and J. Schmidhuber · 2010
Later among the works it cites.
Deep auto-encoder neural networks in reinforcement learning
S. Lange and M. Riedmiller · 2010
Later among the works it cites.
Reinforcement learning on slow features of high-dimensional input streams
R. Legenstein, N. Wilbert, and L. Wiskott · 2010
Later among the works it cites.
GQ( λ \lambda ): A general gradient algorithm for temporal-difference prediction learning with eligibility traces
H. R. Maei and R. S. Sutton · 2010
Later among the works it cites.
Free-energy-based reinforcement learning in a partially observable environment
M. Otsuka, J. Yoshimoto, and K. Doya · 2010
Later among the works it cites.
Policy gradient methods
J. Peters · 2010
Later among the works it cites.
Metalearning
T. Schaul and J. Schmidhuber · 2010
Later among the works it cites.
Evaluation of pooling operations in convolutional architectures for object recognition
D. Scherer, A. Müller, and S. Behnke · 2010
Later among the works it cites.
Formal theory of creativity, fun, and intrinsic motivation (1990-2010)
J. Schmidhuber · 2010
Later among the works it cites.
Parameter-exploring policy gradients
F. Sehnke, C. Osendorfer, T. Rückstieß, A. Graves, J. Peters, and J. Schmidhuber · 2010
Later among the works it cites.
Recurrent policy gradients
D. Wierstra, A. Foerster, J. Peters, and J. Schmidhuber · 2010
Later among the works it cites.
Convolutional neural network committees for handwritten character classification
D. C. Ciresan, U. Meier, L. M. Gambardella, and J. Schmidhuber · 2011
Later among the works it cites.
Flexible, high performance convolutional neural networks for image classification
D. C. Ciresan, U. Meier, J. Masci, L. M. Gambardella, and J. Schmidhuber · 2011
Later among the works it cites.
A committee of neural networks for traffic sign classification
D. C. Ciresan, U. Meier, J. Masci, and J. Schmidhuber · 2011
Later among the works it cites.
Intrinsically motivated evolutionary search for vision-based reinforcement learning
G. Cuccu, M. Luciw, J. Schmidhuber, and F. Gomez · 2011
Later among the works it cites.
Sequential constant size compressor for reinforcement learning
L. Gisslen, M. Luciw, V. Graziano, and J. Schmidhuber · 2011
Later among the works it cites.
Learning recurrent neural networks with Hessian-free optimization
J. Martens and I. Sutskever · 2011
Later among the works it cites.
Better digit recognition with a committee of simple neural nets
U. Meier, D. C. Ciresan, L. M. Gambardella, and J. Schmidhuber · 2011
Later among the works it cites.
The two-dimensional organization of behavior
M. Ring, T. Schaul, and J. Schmidhuber · 2011
Later among the works it cites.
On fast deep nets for AGI vision
J. Schmidhuber, D. Ciresan, U. Meier, J. Masci, and A. Graves · 2011
Later among the works it cites.
A survey of monte carlo tree search methods
C. B. Browne, E. Powley, D. Whitehouse, S. M. Lucas, P. I. Cowling, P. Rohlfshagen, S. Tavener, D. Perez, S. Samothrakis, and S. Colton · 2012
Later among the works it cites.
Deep neural networks segment neuronal membranes in electron microscopy images
D. C. Ciresan, A. Giusti, L. M. Gambardella, and J. Schmidhuber · 2012
Later among the works it cites.
Multi-column deep neural networks for image classification
D. C. Ciresan, U. Meier, and J. Schmidhuber · 2012
Later among the works it cites.
Transfer learning for Latin and Chinese characters with deep neural networks
D. C. Ciresan, U. Meier, and J. Schmidhuber · 2012
Later among the works it cites.
A survey of actor-critic reinforcement learning: Standard and natural policy gradients
I. Grondman, L. Busoniu, G. A. D. Lopes, and R. Babuska · 2012
Later among the works it cites.
Actor-critic reinforcement learning with energy-based policies
N. Heess, D. Silver, and Y. W. Teh · 2012
Later among the works it cites.
Incremental slow feature analysis: Adaptive low-complexity slow feature updating from high-dimensional input streams
V. R. Kompella, M. D. Luciw, and J. Schmidhuber · 2012
Later among the works it cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Later among the works it cites.
Autonomous reinforcement learning on raw visual input data in a real world application
M. Riedmiller, S. Lange, and A. Voigtlaender · 2012
Later among the works it cites.
Self-delimiting neural networks
J. Schmidhuber · 2012
Later among the works it cites.
Reinforcement learning in continuous state and action spaces
H. van Hasselt · 2012
Later among the works it cites.
Evolutionary computation for reinforcement learning
S. Whiteson · 2012
Later among the works it cites.
Reinforcement Learning
M. Wiering and M. van Otterlo · 2012
Later among the works it cites.
Forecasting with recurrent neural networks: 12 tricks
H.-G. Zimmermann, C. Tietz, and R. Grothmann · 2012
Later among the works it cites.
Representation learning: A review and new perspectives
Y. Bengio, A. Courville, and P. Vincent · 2013
Later among the works it cites.
Mitosis detection in breast cancer histology images with deep neural networks
D. C. Ciresan, A. Giusti, L. M. Gambardella, and J. Schmidhuber · 2013
Later among the works it cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2013
Later among the works it cites.
Maxout networks
I. J. Goodfellow, D. Warde-Farley, M. Mirza, A. Courville, and Y. Bengio · 2013
Later among the works it cites.
Evolving large-scale neural networks for vision-based reinforcement learning
J. Koutník, G. Cuccu, J. Schmidhuber, and F. Gomez · 2013
Later among the works it cites.
An intrinsic value system for developing multiple invariant representations with incremental slowness learning
M. Luciw, V. R. Kompella, S. Kazerounian, and J. Schmidhuber · 2013
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2013
Later among the works it cites.
Intrinsically motivated learning of real world sensorimotor skills with developmental constraints
P.-Y. Oudeyer, A. Baranes, and F. Kaplan · 2013
Later among the works it cites.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Later among the works it cites.
PowerPlay
J. Schmidhuber · 2013
Later among the works it cites.
Compete to compute
R. K. Srivastava, J. Masci, S. Kazerounian, F. Gomez, and J. Schmidhuber · 2013
Later among the works it cites.
First experiments with PowerPlay
R. K. Srivastava, B. R. Steunebrink, and J. Schmidhuber · 2013
Later among the works it cites.
A Linear Time Natural Evolution Strategy for Non-Separable Functions
Y. Sun, F. Gomez, T. Schaul, and J. Schmidhuber · 2013
Later among the works it cites.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2013
Later among the works it cites.
TTS synthesis with bidirectional LSTM based recurrent neural networks
Y. Fan, Y. Qian, F. Xie, and F. K. Soong · 2014
Later among the works it cites.
Prosody contour prediction with Long Short-Term Memory, bi-directional, deep recurrent neural networks
R. Fernandez, A. Rendel, B. Ramabhadran, and R. Hoory · 2014
Later among the works it cites.
Multi-digit number recognition from street view imagery using deep convolutional neural networks
I. J. Goodfellow, Y. Bulatov, J. Ibarz, S. Arnoud, and V. Shet · 2014
Later among the works it cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Later among the works it cites.
Deep learning for real-time Atari game play using offline Monte-Carlo tree search planning
X. Guo, S. Singh, H. Lee, R. Lewis, and X. Wang · 2014
Later among the works it cites.
DeepSpeech: Scaling up end-to-end speech recognition
A. Hannun, C. Case, J. Casper, B. Catanzaro, G. Diamos, E. Elsen, R. Prenger, S. Satheesh, S. Sengupta, A. Coates, and A. Y. Ng · 2014
Later among the works it cites.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Later among the works it cites.
J. Koutník, K. Greff, F. Gomez, and J. Schmidhuber · 2014
Later among the works it cites.
Long Short-Term Memory recurrent neural network architectures for large scale acoustic modeling
H. Sak, A. Senior, and F. Beaufays · 2014
Later among the works it cites.
Deep learning in neural networks: An overview
J. Schmidhuber · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Later among the works it cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2014
Later among the works it cites.
Show and tell: A neural image caption generator
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan · 2014
Later among the works it cites.
J. Weston, S. Chopra, and A. Bordes · 2014
Later among the works it cites.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Closest in time.
Google Voice search: faster and more accurate
H. Sak, A. Senior, K. Rao, F. Beaufays, and J. Schalkwyk · 2015
Closest in time.
Deep Learning
J. Schmidhuber · 2015
Closest in time.