Fetching the paper…
Reading the bibliography…
While neural networks have acted as a strong unifying force in the design of modern AI systems, the neural network architectures themselves remain highly heterogeneous due to the variety of tasks to be solved.
1909
Earlier work this paper cites.
Munro, P.: A dual back-propagation scheme for scalar reward learning. In: Proceedings of the Ninth Annual Conference of the Cognitive Science Society. pp. 165–176 (1987)
1987
Earlier work this paper cites.
Robinson, A.J.: Dynamic error propagation networks. Ph.D. thesis, Trinity Hall and Cambridge University Engineering Department (1989)
1989
Earlier work this paper cites.
Robinson, T., Fallside, F.: Dynamic reinforcement driven error propagation networks with application to game playing. In: Proceedings of the 11th Conference of the Cognitive Science Society, Ann Arbor. pp. 836–843 (1989)
1989
Earlier work this paper cites.
Hochreiter, S.: Implementierung und Anwendung eines ‘neuronalen’ Echtzeit-Lernalgorithmus für reaktive Umgebungen. Practical work, Institut für Informatik, Technische Universität München (1990)
1990
Earlier work this paper cites.
Schmidhuber, J.: Making the world differentiable: On using fully recurrent self-supervised neural networks for dynamic reinforcement learning and planning in non-stationary environments. Tech. Rep. FKI-126-90 (revised), Institut für Informatik, Technische Universität München (1990), Experiments by Sepp Hochreiter
1990
Earlier work this paper cites.
Hochreiter, S.: Untersuchungen zu dynamischen neuronalen Netzen. Master’s thesis. Institut für Informatik, Technische Universität München (1991)
1991
Earlier work this paper cites.
Bengio, Y., Simard, P., Frasconi, P.: Learning long-term dependencies with gradient descent is difficult. IEEE Transactions on Neural Networks 5
1994
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J.: Long short-term memory. Tech. Rep. FKI-207-95, Fakultät für Informatik, Technische Universität München (1995)
1995
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J.: LSTM can solve hard long time lag problems. In: Advances in Neural Information Processing Systems 9 (NIPS). pp. 473–479 (1996)
1996
Earlier work this paper cites.
Hochreiter, S.: Recurrent neural net learning and vanishing gradient. In: Freksa, C. (ed.) Proceedings in Artificial Intelligence - Fuzzy-Neuro-Systeme’97 Workshop. pp. 130–137. Infix (1997)
1997
Earlier work this paper cites.
Hochreiter, S., Schmidhuber, J.: Long short-term memory. Neural Computation 9
1997
Earlier work this paper cites.
Hochreiter, S.: The vanishing gradient problem during learning recurrent neural nets and problem solutions. International Journal of Uncertainty, Fuzziness and Knowledge-Based Systems 6
1998
Earlier work this paper cites.
Gers, F.A., Schmidhuber, J., Cummins, F.: Learning to forget: Continual prediction with LSTM. In: Proceedings of the International Conference on Artificial Neural Networks (ICANN). vol. 2, pp. 850–855 (1999)
1999
Earlier work this paper cites.
Gers, F.A., Schmidhuber, J.: Recurrent nets that time and count. In: Proceedings of the IEEE International Joint Conference on Neural Networks (IJCNN). vol. 3, pp. 189–194 (2000)
2000
Earlier work this paper cites.
Gers, F.A., Schmidhuber, J., Cummins, F.: Learning to forget: Continual prediction with LSTM. Neural Computation 12
2000
Earlier work this paper cites.
Hochreiter, S., Bengio, Y., Frasconi, P., Schmidhuber, J.: Gradient flow in recurrent nets: the difficulty of learning long-term dependencies. In: Kolen, J.F., Kremer, S.C. (eds.) A Field Guide to Dynamical Recurrent Networks, pp. 237–244. IEEE Press (2001)
2001
Earlier work this paper cites.
Hochreiter, S., Younger, A.S., Conwell, P.R.: Learning to learn using gradient descent. In: Proceedings of the International Conference on Artificial Neural Networks (ICANN). pp. 87–94 (2001)
2001
Earlier work this paper cites.
Bakker, B.: Reinforcement learning with long short-term memory. In: Advances in Neural Information Processing Systems 14 (NIPS). pp. 1475–1482 (2002)
2002
Earlier work this paper cites.
Gevrey, M., Dimopoulos, I., Lek, S.: Review and comparison of methods to study the contribution of variables in artificial neural network models. Ecological Modelling 160
2003
Earlier work this paper cites.
Graves, A., Schmidhuber, J.: Framewise phoneme classification with bidirectional LSTM and other neural network architectures. Neural Networks 18
2005
Earlier work this paper cites.
Bakker, B.: Reinforcement learning by backpropagation through an LSTM model/critic. In: IEEE International Symposium on Approximate Dynamic Programming and Reinforcement Learning. pp. 127–134 (2007)
2007
Earlier work this paper cites.
Hochreiter, S., Heusel, M., Obermayer, K.: Fast model-based protein homology detection without alignment. Bioinformatics 23
2007
Earlier work this paper cites.
Graves, A., Liwicki, M., Fernandez, S., Bertolami, R., Bunke, H., Schmidhuber, J.: A novel connectionist system for unconstrained handwriting recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence 31
2009
Earlier work this paper cites.
Rahmandad, H., Repenning, N., Sterman, J.: Effects of feedback delay on learning. System Dynamics Review 25
2009
Earlier work this paper cites.
Landecker, W., Thomure, M.D., Bettencourt, L.M.A., Mitchell, M., Kenyon, G.T., Brumby, S.P.: Interpreting individual classifications of hierarchical networks. In: IEEE Symposium on Computational Intelligence and Data Mining (CIDM). pp. 32–38 (2013)
2013
Earlier work this paper cites.
Socher, R., Perelygin, A., Wu, J.Y., Chuang, J., Manning, C.D., Ng, A.Y., Potts, C.: Recursive deep models for semantic compositionality over a sentiment treebank. In: Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing (EMNLP). pp. 1631–1642. Association for Computational Linguistics (2013)
2013
Earlier work this paper cites.
Cho, K., van Merrienboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., Bengio, Y.: Learning phrase representations using RNN encoder-decoder for statistical machine translation. In: Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP). pp. 1724–1734. Association for Computational Linguistics (2014)
2014
Cited alongside, same era.
Geiger, J.T., Zhang, Z., Weninger, F., Schuller, B., Rigoll, G.: Robust speech recognition using long short-term memory recurrent neural networks for hybrid acoustic modelling. In: Proceedings of the 15th Annual Conference of the International Speech Communication Association (INTERSPEECH). pp. 631–635 (2014)
2014
Cited alongside, same era.
Gonzalez-Dominguez, J., Lopez-Moreno, I., Sak, H., Gonzalez-Rodriguez, J., Moreno, P.J.: Automatic language identification using long short-term memory recurrent neural networks. In: Proceedings of the 15th Annual Conference of the International Speech Communication Association (INTERSPEECH). pp. 2155–2159 (2014)
2014
Cited alongside, same era.
Lapuschkin, S., Binder, A., Müller, K.R., Samek, W.: Understanding and comparing deep neural networks for age and gender classification. In: IEEE International Conference on Computer Vision Workshops. pp. 1629–1638 (2017)
2017
Later among the works it cites.
2017
Later among the works it cites.
Lundberg, S.M., Lee, S.I.: A unified approach to interpreting model predictions. In: Advances in Neural Information Processing Systems 30 (NIPS). pp. 4765–4774 (2017)
2017
Later among the works it cites.
Luoma, J., Ruutu, S., King, A.W., Tikkanen, H.: Time delays, competitive interdependence, and firm performance. Strategic Management Journal 38
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
Marchi, E., Ferroni, G., Eyben, F., Gabrielli, L., Squartini, S., Schuller, B.: Multi-resolution linear prediction based features for audio onset detection with bidirectional LSTM neural networks. In: IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 2164–2168 (2014)
2014
Cited alongside, same era.
Sak, H., Senior, A., Beaufays, F.: Long short-term memory recurrent neural network architectures for large scale acoustic modeling. In: Proceedings of the 15th Annual Conference of the International Speech Communication Association (INTERSPEECH). pp. 338–342. Singapore (2014)
2014
Cited alongside, same era.
Simonyan, K., Vedaldi, A., Zisserman, A.: Deep inside convolutional networks: Visualising image classification models and saliency maps. In: International Conference on Learning Representations (ICLR) (2014)
2014
Cited alongside, same era.
Sutskever, I., Vinyals, O., Le, Q.V.: Sequence to sequence learning with neural networks. In: Advances in Neural Information Processing Systems 27 (NIPS). pp. 3104–3112 (2014)
2014
Cited alongside, same era.
Zeiler, M.D., Fergus, R.: Visualizing and understanding convolutional networks. In: Proceedings of the European Conference on Computer Vision (ECCV). LNCS, vol. 8689, pp. 818–833. Springer, Cham (2014)
2014
Cited alongside, same era.
Bach, S., Binder, A., Montavon, G., Klauschen, F., Müller, K.R., Samek, W.: On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation. PLoS ONE 10
2015
Cited alongside, same era.
2015
Cited alongside, same era.
Hausknecht, M., Stone, P.: Deep recurrent q-learning for partially observable MDPs. In: AAAI Fall Symposium Series - Sequential Decision Making for Intelligent Agents. pp. 29–37 (2015)
2015
Cited alongside, same era.
Montavon, G., Lapuschkin, S., Binder, A., Samek, W., Müller, K.R.: Explaining nonlinear classification decisions with deep Taylor decomposition. Pattern Recognition 65
2017
Later among the works it cites.
Samek, W., Binder, A., Montavon, G., Lapuschkin, S., Müller, K.R.: Evaluating the visualization of what a deep neural network has learned. IEEE Transactions on Neural Networks and Learning Systems 28
2017
Later among the works it cites.
Shrikumar, A., Greenside, P., Kundaje, A.: Learning important features through propagating activation differences. In: Proceedings of the 34th International Conference on Machine Learning (ICML). vol. 70, pp. 3145–3153 (2017)
2017
Later among the works it cites.
Sundararajan, M., Taly, A., Yan, Q.: Axiomatic attribution for deep networks. In: Proceedings of the 34th International Conference on Machine Learning (ICML). vol. 70, pp. 3319–3328 (2017)
2017
Later among the works it cites.
Sutton, R.S., Barto, A.G.: Reinforcement learning: An introduction. MIT Press, Cambridge, MA, 2 edn. (2017), draft from November 2017
2017
Later among the works it cites.
Ancona, M., Ceolini, E., Öztireli, C., Gross, M.: Towards better understanding of gradient-based attribution methods for deep neural networks. In: International Conference on Learning Representations (ICLR) (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Chen, J., Song, L., Wainwright, M., Jordan, M.: Learning to explain: An information-theoretic perspective on model interpretation. In: Proceedings of the 35th International Conference on Machine Learning (ICML). vol. 80, pp. 883–892 (2018)
2018
Later among the works it cites.
Montavon, G., Samek, W., Müller, K.R.: Methods for interpreting and understanding deep neural networks. Digital Signal Processing 73
2018
Later among the works it cites.
Morcos, A.S., Barrett, D.G., Rabinowitz, N.C., Botvinick, M.: On the importance of single directions for generalization. In: International Conference on Learning Representations (ICLR) (2018)
2018
Later among the works it cites.
Murdoch, W.J., Liu, P.J., Yu, B.: Beyond word importance: Contextual decomposition to extract interactions from LSTMs. In: International Conference on Learning Representations (ICLR) (2018)
2018
Later among the works it cites.
Poerner, N., Schütze, H., Roth, B.: Evaluating neural network explanation methods using hybrid documents and morphosyntactic agreement. In: Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (ACL). pp. 340–350. Association for Computational Linguistics (2018)
2018
Later among the works it cites.
Rieger, L., Chormai, P., Montavon, G., Hansen, L.K., Müller, K.R.: Structuring neural networks for more explainable predictions. In: Escalante H. et al. (eds.) Explainable and Interpretable Models in Computer Vision and Machine Learning. TSSCML, pp. 115–131. Springer, Cham (2018)
2018
Later among the works it cites.
Sahni, H.: Reinforcement learning never worked, and ’deep’ only helped a bit. himanshusahni.github.io/2018/02/23/reinforcement-learning-never-worked.html (2018)
2018
Later among the works it cites.
Thuillier, E., Gamper, H., Tashev, I.J.: Spatial audio feature discovery with convolutional neural networks. In: IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 6797–6801 (2018)
2018
Later among the works it cites.
Yang, Y., Tresp, V., Wunderle, M., Fasching, P.A.: Explaining therapy predictions with layer-wise relevance propagation in neural networks. In: IEEE International Conference on Healthcare Informatics (ICHI). pp. 152–162 (2018)
2018
Later among the works it cites.
Arras, L., Osman, A., Müller, K.R., Samek, W.: Evaluating recurrent neural network explanations. In: Proceedings of the ACL 2019 Workshop on BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP, pp. 113–126. Association for Computational Linguistics (2019)
2019
Closest in time.
Horst, F., Lapuschkin, S., Samek, W., Müller, K.R., Schöllhorn, W.I.: Explaining the unique nature of individual gait patterns with deep learning. Scientific Reports 9
2019
Closest in time.
2019
Closest in time.
Lapuschkin, S., Wäldchen, S., Binder, A., Montavon, G., Samek, W., Müller, K.R.: Unmasking clever hans predictors and assessing what machines really learn. Nature Communications 10
2019
Closest in time.
Montavon, G., Binder, A., Lapuschkin, S., Samek, W., Müller, K.R.: Layer-wise relevance propagation: An overview. In: Explainable AI: Interpreting, Explaining and Visualizing Deep Learning. Lecture Notes in Computer Science 11700
2019
Closest in time.