Fetching the paper…
Reading the bibliography…
This paper presents an empirical study on the weights of neural networks, where we interpret each model as a point in a high-dimensional space -- the neural weight space.
Marcus Gallagher and Tom Downs, ‘Visualization of learning in neural networks using principal component analysis’, in International Conference on Computational Intelligence and Multimedia Applications (ICCIMA 1997)
1997
Earlier work this paper cites.
Marcus Gallagher and Tom Downs, ‘Weight space learning trajectory visualization’, in Australian Conference on Neural Networks (ACNN 1997)
1997
Earlier work this paper cites.
Yann LeCun, Léon Bottou, Yoshua Bengio, Patrick Haffner, et al., ‘Gradient-based learning applied to document recognition’, Proceedings of the IEEE
1998
Earlier work this paper cites.
Alex Krizhevsky and Geoffrey Hinton, ‘Learning multiple layers of features from tiny images’, Technical report, Citeseer, (2009)
2009
Earlier work this paper cites.
Dumitru Erhan, Yoshua Bengio, Aaron Courville, Pierre-Antoine Manzagol, Pascal Vincent, and Samy Bengio, ‘Why does unsupervised pre-training help deep learning?’, Journal of Machine Learning Research (JMLR)
2010
Earlier work this paper cites.
Xavier Glorot and Yoshua Bengio, ‘Understanding the difficulty of training deep feedforward neural networks’, in International conference on artificial intelligence and statistics (AISTATS 2010)
2010
Earlier work this paper cites.
Vinod Nair and Geoffrey E Hinton, ‘Rectified linear units improve restricted boltzmann machines’, in International conference on machine learning (ICML 2010)
2010
Earlier work this paper cites.
Adam Coates, Andrew Ng, and Honglak Lee, ‘An analysis of single-layer networks in unsupervised feature learning’, in International conference on artificial intelligence and statistics (AISTATS 2011)
2011
Earlier work this paper cites.
Frank Hutter, Holger H Hoos, and Kevin Leyton-Brown, ‘Sequential model-based optimization for general algorithm configuration’, in International Conference on Learning and Intelligent Optimization (LION 2011)
2011
Earlier work this paper cites.
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y Ng, ‘Reading digits in natural images with unsupervised feature learning’, in NIPS Workshop on Deep Learning and Unsupervised Feature Learning
2011
Earlier work this paper cites.
James Bergstra and Yoshua Bengio, ‘Random search for hyper-parameter optimization’, Journal of Machine Learning Research (JMLR)
2012
Earlier work this paper cites.
Geoffrey Hinton, Nitish Srivastava, and Kevin Swersky, ‘Neural networks for machine learning lecture 6a overview of mini-batch gradient descent’, (2012)
2012
Earlier work this paper cites.
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton, ‘Imagenet classification with deep convolutional neural networks’, in Advances in neural information processing systems (NIPS 2012)
2012
Earlier work this paper cites.
Jonas Mockus, Bayesian approach to global optimization: theory and applications
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, ‘Generative adversarial nets’, in International Conference on Neural Information Processing Systems (NIPS 2014)
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov, ‘Dropout: a simple way to prevent neural networks from overfitting’, The Journal of Machine Learning Research (JMLR)
2014
Cited alongside, same era.
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson, ‘How transferable are features in deep neural networks?’, in International Conference on Neural Information Processing Systems (NIPS 2014)
2014
Cited alongside, same era.
Matthew D Zeiler and Rob Fergus, ‘Visualizing and understanding convolutional networks’, in European conference on computer vision (ECCV 2014)
2014
Cited alongside, same era.
Giuseppe Ateniese, Luigi V Mancini, Angelo Spognardi, Antonio Villani, Domenico Vitali, and Giovanni Felici, ‘Hacking smart machines with smarter ones: How to extract meaningful data from machine learning classifiers’, International Journal of Security and Networks (IJSN)
2015
Cited alongside, same era.
Maithra Raghu, Justin Gilmer, Jason Yosinski, and Jascha Sohl-Dickstein, ‘Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability’, in International Conference on Neural Information Processing Systems (NIPS 2017)
2017
Later among the works it cites.
Paulo E Rauber, Samuel G Fadel, Alexandre X Falcao, and Alexandru C Telea, ‘Visualizing the hidden activity of artificial neural networks’, IEEE transactions on visualization and computer graphics (TVCG)
2017
Later among the works it cites.
Esteban Real, Sherry Moore, Andrew Selle, Saurabh Saxena, Yutaka Leon Suematsu, Jie Tan, Quoc V Le, and Alexey Kurakin, ‘Large-scale evolution of image classifiers’, in International Conference on Machine Learning (ICML 2017)
2017
Later among the works it cites.
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra, ‘Grad-cam: Visual explanations from deep networks via gradient-based localization’, in IEEE International Conference on Computer Vision (CVPR 2017)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
Matt Fredrikson, Somesh Jha, and Thomas Ristenpart, ‘Model inversion attacks that exploit confidence information and basic countermeasures’, in ACM SIGSAC Conference on Computer and Communications Security (CCS 2015)
2015
Cited alongside, same era.
Sergey Ioffe and Christian Szegedy, ‘Batch normalization: Accelerating deep network training by reducing internal covariate shift’, in International Conference on Machine Learning (ICML 2015)
2015
Cited alongside, same era.
Yixuan Li, Jason Yosinski, Jeff Clune, Hod Lipson, and John Hopcroft, ‘Convergent learning: Do different neural networks learn the same representations?’, in NIPS Workshop on Feature Extraction: Modern Questions and Challenges
2015
Cited alongside, same era.
Aravindh Mahendran and Andrea Vedaldi, ‘Understanding deep image representations by inverting them’, in IEEE conference on computer vision and pattern recognition (CVPR 2015)
2015
Cited alongside, same era.
Jason Yosinski, Jeff Clune, Anh Nguyen, Thomas Fuchs, and Hod Lipson, ‘Understanding neural networks through deep visualization’, in ICML Workshop on Deep Learning
2015
Cited alongside, same era.
Bolei Zhou, Aditya Khosla, Agata Lapedriza, Aude Oliva, and Antonio Torralba, ‘Object detectors emerge in deep scene CNNs’, in International Conference on Learning Representations (ICLR 2015)
2015
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Later among the works it cites.
Reza Shokri, Marco Stronati, Congzheng Song, and Vitaly Shmatikov, ‘Membership inference attacks against machine learning models’, in IEEE Symposium on Security and Privacy (SP)
2017
Later among the works it cites.
2017
Later among the works it cites.
Barret Zoph and Quoc V. Le, ‘Neural architecture search with reinforcement learning’, in International Conference on Learning Representations (ICLR 2017)
2017
Later among the works it cites.
Joseph Antognini and Jascha Sohl-Dickstein, ‘PCA of high dimensional random walks with comparison to neural network training’, in Advances in Neural Information Processing Systems (NeurIPS 2018)
2018
Later among the works it cites.
Alsallakh Bilal, Amin Jourabloo, Mao Ye, Xiaoming Liu, and Liu Ren, ‘Do convolutional neural networks learn class hierarchy?’, IEEE transactions on visualization and computer graphics (TVCG)
2018
Later among the works it cites.
Fred Matthew Hohman, Minsuk Kahng, Robert Pienta, and Duen Horng Chau, ‘Visual analytics in deep learning: An interrogative survey for the next frontiers’, IEEE transactions on visualization and computer graphics (TVCG)
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
Roman Novak, Yasaman Bahri, Daniel A. Abolafia, Jeffrey Pennington, and Jascha Sohl-Dickstein, ‘Sensitivity and generalization in neural networks: an empirical study’, in International Conference on Learning Representations (ICLR 2018)
2018
Later among the works it cites.
Amir R Zamir, Alexander Sax, William Shen, Leonidas J Guibas, Jitendra Malik, and Silvio Savarese, ‘Taskonomy: Disentangling task transfer learning’, in IEEE Conference on Computer Vision and Pattern Recognition (CVPR 2018)
2018
Later among the works it cites.
Bolei Zhou, David Bau, Aude Oliva, and Antonio Torralba, ‘Interpreting deep visual representations via network dissection’, IEEE transactions on pattern analysis and machine intelligence (TPAMI)
2018
Later among the works it cites.
Barret Zoph, Vijay Vasudevan, Jonathon Shlens, and Quoc V Le, ‘Learning transferable architectures for scalable image recognition’, in IEEE conference on computer vision and pattern recognition (CVPR 2018)
2018
Later among the works it cites.
2019
Later among the works it cites.
Hanxiao Liu, Karen Simonyan, and Yiming Yang, ‘DARTS: Differentiable architecture search’, in International Conference on Learning Representations (ICLR 2019)
2019
Later among the works it cites.