Fetching the paper…
Reading the bibliography…
Neural networks are usually not the tool of choice for nonparametric high-dimensional problems where the number of input features is much larger than the number of observations.
Nadaraya, E. A. (1964), ‘On estimating regression’, Theory of Probability & Its Applications
1964
Earlier work this paper cites.
Watson, G. S. (1964), ‘Smooth regression analysis’, Sankhyā: The Indian Journal of Statistics, Series A
1964
Earlier work this paper cites.
Nelder, J. A. & Mead, R. (1965), ‘A simplex method for function minimization’, The computer journal
1965
Earlier work this paper cites.
Breiman, L., Friedman, J., Stone, C. J. & Olshen, R. A. (1984), Classification and regression trees
1984
Earlier work this paper cites.
Kidera, A., Konishi, Y., Oka, M., Ooi, T. & Scheraga, H. A. (1985), ‘Statistical analysis of the physical properties of the 20 naturally occurring amino acids’, J. Protein Chem
1985
Earlier work this paper cites.
Barzilai, J. & Borwein, J. M. (1988), ‘Two-point step size gradient methods’, IMA journal of numerical analysis
1988
Earlier work this paper cites.
Barron, A. R. (1993), ‘Universal approximation bounds for superpositions of a sigmoidal function’, IEEE Transactions on Information theory
1993
Earlier work this paper cites.
Leshno, M., Lin, V. Y., Pinkus, A. & Schocken, S. (1993), ‘Multilayer feedforward networks with a nonpolynomial activation function can approximate any function’, Neural Netw
1993
Earlier work this paper cites.
Fefferman, C. (1994), ‘Reconstructing a neural net from its output’, Revista Matemática Iberoamericana
1994
Earlier work this paper cites.
Tibshirani, R. (1996), ‘Regression shrinkage and selection via the lasso’, Journal of the Royal Statistical Society. Series B (Methodological)
1996
Earlier work this paper cites.
van der Vaart, A. W. & Wellner, J. A. (1996), Weak convergence and empirical processes: with applications to statistics
1996
Earlier work this paper cites.
Bartlett, P. L. (1998), ‘The sample complexity of pattern classification with neural networks: the size of the weights is more important than the size of the network’, IEEE transactions on Information Theory
1998
Earlier work this paper cites.
Sun, X. (1999), The Lasso and its implementation for neural networks, PhD thesis, National Library of Canada= Bibliothèque nationale du Canada
1999
Earlier work this paper cites.
Cristianini, N. & Shawe-Taylor, J. (2000), An introduction to support vector machines and other kernel-based learning methods
2000
Earlier work this paper cites.
van de Geer, S. A. (2000), Empirical Processes in M-estimation
2000
Earlier work this paper cites.
Breiman, L. (2001), ‘Random forests’, Mach. Learn
2001
Earlier work this paper cites.
Chiaretti, S., Li, X., Gentleman, R., Vitale, A., Vignetti, M., Mandelli, F., Ritz, J. & Foa, R. (2004), ‘Gene expression profile of adult T-cell acute lymphocytic leukemia identifies distinct subsets of patients with different response to therapy and survival’, Blood
2004
Earlier work this paper cites.
Daubechies, I., Defrise, M. & De Mol, C. (2004), ‘An iterative thresholding algorithm for linear inverse problems with a sparsity constraint’, Communications on pure and applied mathematics
2004
Cited alongside, same era.
Nesterov, Y. (2004), Introductory lectures on convex optimization: A basic course
2004
Cited alongside, same era.
Yuan, M. & Lin, Y. (2006), ‘Model selection and estimation in regression with grouped variables’, Journal of the Royal Statistical Society: Series B (Statistical Methodology)
2006
Cited alongside, same era.
Ravikumar, P., Liu, H., Lafferty, J. & Wasserman, L. (2007), Spam: Sparse additive models, in
2007
Cited alongside, same era.
Hofmann, T., Schölkopf, B. & Smola, A. J. (2008), ‘Kernel methods in machine learning’, The Annals of Statistics
2008
Cited alongside, same era.
Bühlmann, P., Kalisch, M. & Meier, L. (2014), ‘High-dimensional statistics with a view toward applications in biology’, Annual Review of Statistics and Its Application
2014
Later among the works it cites.
Kim, Y., Sidney, J., Buus, S., Sette, A., Nielsen, M. & Peters, B. (2014), ‘Dataset size and composition impact the reliability of performance benchmarks for peptide-MHC binding predictions’, BMC Bioinformatics
2014
Later among the works it cites.
Collins, F. S. & Varmus, H. (2015), ‘A new initiative on precision medicine’, N. Engl. J. Med
2015
Later among the works it cites.
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V. & Rabinovich, A. (2015), Going deeper with convolutions, in
2015
Later among the works it cites.
Trolle, T., Metushi, I. G., Greenbaum, J. A., Kim, Y., Sidney, J., Lund, O., Sette, A., Peters, B. & Nielsen, M. (2015), ‘Automated benchmarking of peptide-MHC class I binding predictions’, Bioinformatics
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Anthony, M. & Bartlett, P. L. (2009), Neural Network Learning: Theoretical Foundations
2009
Cited alongside, same era.
Beck, A. & Teboulle, M. (2009), ‘A fast iterative shrinkage-thresholding algorithm for linear inverse problems’, SIAM journal on imaging sciences
2009
Cited alongside, same era.
Meier, L., Van de Geer, S., Bühlmann, P. et al. (2009), ‘High-dimensional additive modeling’, The Annals of Statistics
2009
Cited alongside, same era.
Städler, N., Bühlmann, P. & Van De Geer, S. (2010), ‘L1-penalization for mixture regression models’, Test
2010
Cited alongside, same era.
Bühlmann, P. & Van De Geer, S. (2011), Statistics for high-dimensional data: methods, theory and applications
2011
Cited alongside, same era.
Krizhevsky, A., Sutskever, I. & Hinton, G. E. (2012), Imagenet classification with deep convolutional neural networks, in
2012
Cited alongside, same era.
Snoek, J., Larochelle, H. & Adams, R. P. (2012), Practical bayesian optimization of machine learning algorithms, in
2012
Cited alongside, same era.
2015
Later among the works it cites.
Vita, R., Overton, J. A., Greenbaum, J. A., Ponomarenko, J., Clark, J. D., Cantrell, J. R., Wheeler, D. K., Gabbard, J. L., Hix, D., Sette, A. & Peters, B. (2015), ‘The immune epitope database (IEDB) 3.0’, Nucleic Acids Res
2015
Later among the works it cites.
Alvarez, J. M. & Salzmann, M. (2016), Learning the number of neurons in deep networks, in
2016
Later among the works it cites.
Andreatta, M. & Nielsen, M. (2016), ‘Gapped sequence alignment using artificial neural networks: application to the MHC class I system’, Bioinformatics
2016
Later among the works it cites.
Ghadimi, S. & Lan, G. (2016), ‘Accelerated gradient methods for nonconvex nonlinear and stochastic programming’, Mathematical Programming
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
Bach, F. (2017), ‘Breaking the curse of dimensionality with convex neural networks’, Journal of Machine Learning Research
2017
Closest in time.
Jurtz, V., Paul, S., Andreatta, M., Marcatili, P., Peters, B. & Nielsen, M. (2017), ‘NetMHCpan-4.0: Improved Peptide-MHC class I interaction predictions integrating eluted ligand and peptide binding affinity data’, J. Immunol
2017
Closest in time.
Boehm, K. M., Bhinder, B., Raja, V. J., Dephoure, N. & Elemento, O. (2018), Predicting peptide presentation by major histocompatibility complex class I using one million peptides
2018
Closest in time.
Li, X. (2018), ‘All: A data package’, R package version 1.24.0
2018
Closest in time.
O’Donnell, T. J., Rubinsteyn, A., Bonsack, M., Riemer, A. B., Laserson, U. & Hammerbacher, J. (2018), ‘MHCflurry: Open-Source class I MHC binding affinity prediction’, Cell Syst
2018
Closest in time.