Fetching the paper…
Reading the bibliography…
Jacobian-vector products (JVPs) form the backbone of many recent developments in Deep Networks (DNs), with applications including faster constrained optimization, regularization with generalization guarantees, and adversarial example sensitivity assessments.
Calculating the singular values and pseudo-inverse of a matrix
Golub, G. and Kahan, W · 1965
Earlier work this paper cites.
Differential equations, dynamical systems, and linear algebra (pure and applied mathematics, vol. 60)
Hirsch, M. and Smale, S · 1974
Earlier work this paper cites.
Taylor expansion of the accumulated rounding error
Linnainmaa, S · 1976
Earlier work this paper cites.
A practical guide to splines , volume 27
De Boor, C., De Boor, C., Mathématicien, E.-U., De Boor, C., and De Boor, C · 1978
Earlier work this paper cites.
Sample-based non-uniform random variate generation
Devroye, L · 1986
Earlier work this paper cites.
Learning representations by back-propagating errors
Rumelhart, D. E., Hinton, G. E., and Williams, R. J · 1986
Earlier work this paper cites.
Small-sample statistical estimates for matrix norms
Gudmundsson, T., Kenney, C. S., and Laub, A. J · 1995
Earlier work this paper cites.
Matrix computations johns hopkins university press
Golub, G. H. and Van Loan, C. F · 1996
Earlier work this paper cites.
Face recognition: A convolutional neural-network approach
Lawrence, S., Giles, C. L., Tsoi, A. C., and Back, A. D · 1997
Earlier work this paper cites.
Numerical linear algebra , volume 50
Trefethen, L. N. and Bau III, D · 1997
Earlier work this paper cites.
The idea behind krylov methods
Ipsen, I. C. F. and Meyer, C. D · 1998
Earlier work this paper cites.
Statistical condition estimation for linear systems
Kenney, C. S., Laub, A. J., and Reese, M · 1998
Earlier work this paper cites.
Radial and uniform distributions in vector and matrix spaces for probabilistic robustness
Calafiore, G., Dabbene, F., and Tempo, R · 1999
Earlier work this paper cites.
Matrix analysis and applied linear algebra , volume 71
Meyer, C. D · 2000
Earlier work this paper cites.
Spline functions: basic theory
Schumaker, L · 2007
Earlier work this paper cites.
Evaluating derivatives: principles and techniques of algorithmic differentiation
Griewank, A. and Walther, A · 2008
Earlier work this paper cites.
Optimal jacobian accumulation is np-complete
Naumann, U · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Spline functions and multivariate interpolations , volume 248
Bojanov, B. D., Hakopian, H., and Sahakian, B · 2013
Earlier work this paper cites.
Hybrid speech recognition with deep bidirectional lstm
Graves, A., Jaitly, N., and Mohamed, A.-r · 2013
Earlier work this paper cites.
Lin, M., Chen, Q., and Yan, S · 2013
Earlier work this paper cites.
On the number of response regions of deep feed forward networks with piece-wise linear activations
Pascanu, R., Montufar, G., and Bengio, Y · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. and Ba, J · 2014
Earlier work this paper cites.
On the number of linear regions of deep neural networks
Montufar, G. F., Pascanu, R., Cho, K., and Bengio, Y · 2014
Earlier work this paper cites.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Abadi, M., Agarwal, A., Barham, P., Brevdo, E., Chen, Z., Citro, C., Corrado, G. S., Davis, A., Dean, J., Devin, M., Ghemawat, S., Goodfellow, I., Harp, A., Irving, G., Isard, M., Jia, Y., Jozefowicz, R., Kaiser, L., Kudlur, M., Levenberg, J., Mané, D., Monga, R., Moore, S., Murray, D., Olah, C., Schuster, M., Shlens, J., Steiner, B., Sutskever, I., Talwar, K., Tucker, P., Vanhoucke, V., Vasudevan, V., Viégas, F., Vinyals, O., Warden, P., Wattenberg, M., Wicke, M., Yu, Y., and Zheng, X · 2015
Earlier work this paper cites.
Block power method for svd decomposition
Bentbib, A. and Kanber, A · 2015
Cited alongside, same era.
Deep learning
LeCun, Y., Bengio, Y., and Hinton, G · 2015
Cited alongside, same era.
Understanding deep neural networks with rectified linear units
Arora, R., Basu, A., Mianjy, P., and Mukherjee, A · 2016
Cited alongside, same era.
Group equivariant convolutional networks
Cohen, T. and Welling, M · 2016
Cited alongside, same era.
Dropout as a bayesian approximation: Representing model uncertainty in deep learning
Gal, Y. and Ghahramani, Z · 2016
Cited alongside, same era.
Deep Learning , volume 1
Goodfellow, I., Bengio, Y., and Courville, A · 2016
Cited alongside, same era.
Deep learning for self-driving cars: chances and challenges
Rao, Q. and Frtunikj, J · 2018
Later among the works it cites.
Deep learning detecting fraud in credit card transactions
Roy, A., Sun, J., Mahoney, R., Alonzi, L., Adams, S., and Beling, P · 2018
Later among the works it cites.
Bounding and counting linear regions of deep neural networks
Serra, T., Tjandraatmadja, C., and Ramalingam, S · 2018
Later among the works it cites.
Generative adversarial network based telecom fraud detection at the receiving bank
Zheng, Y.-J., Zhou, X.-H., Sheng, W.-G., Xue, Y., and Chen, S.-Y · 2018
Later among the works it cites.
The geometry of deep networks: Power diagram subdivision
Balestriero, R., Cosentino, R., Aazhang, B., and Baraniuk, R · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zagoruyko, S. and Komodakis, N · 2016
Cited alongside, same era.
Brein memory: A single-chip binary/ternary reconfigurable in-memory deep neural network accelerator achieving 1.4 tops at 0.6 w
Ando, K., Ueyoshi, K., Orimo, K., Yonekawa, H., Sato, S., Nakahara, H., Takamaeda-Yamazaki, S., Ikebe, M., Asai, T., Kuroda, T., et al · 2017
Cited alongside, same era.
Automatic differentiation in machine learning: a survey
Baydin, A. G., Pearlmutter, B. A., Radul, A. A., and Siskind, J. M · 2017
Cited alongside, same era.
Adabatch: Adaptive batch sizes for training deep neural networks
Devarakonda, A., Naumov, M., and Garland, M · 2017
Cited alongside, same era.
generating forward-mode graphs by calling tf.gradients twice
Jamie, T., Alexandre, P., and Matthew, J · 2017
Cited alongside, same era.
Imposing hard constraints on deep networks: Promises and limitations
Márquez-Neila, P., Salzmann, M., and Fua, P · 2017
Cited alongside, same era.
Bartlett, P. L., Harvey, N., Liaw, C., and Mehrabian, A · 2019
Later among the works it cites.
Neural architecture search: A survey
Elsken, T., Metzen, J. H., Hutter, F., et al · 2019
Later among the works it cites.
Complexity of linear regions in deep networks
Hanin, B. and Rolnick, D · 2019
Later among the works it cites.
Orthogonal deep neural networks
Jia, K., Li, S., Wen, Y., Liu, T., and Tao, D · 2019
Later among the works it cites.
A video-based fire detection using deep learning models
Kim, B. and Lee, J · 2019
Later among the works it cites.
Implicit rugosity regularization via data augmentation
LeJeune, D., Balestriero, R., Javadi, H., and Baraniuk, R. G · 2019
Later among the works it cites.
Dissecting deep neural networks
Robinson, H., Rasheed, A., and San, O · 2019
Later among the works it cites.
A survey on image data augmentation for deep learning
Shorten, C. and Khoshgoftaar, T. M · 2019
Later among the works it cites.
Linear algebra and learning from data
Strang, G · 2019
Later among the works it cites.
A representer theorem for deep neural networks
Unser, M · 2019
Later among the works it cites.
Mad max: Affine spline insights into deep learning
Balestriero, R. and Baraniuk, R. G · 2020
Later among the works it cites.
Max-affine spline insights into deep generative networks
Balestriero, R., Paris, S., and Baraniuk, R · 2020
Later among the works it cites.
Provable finite data generalization with group autoencoder, 2020
Cosentino, R., Balestriero, R., Baraniuk, R., and Aazhang, B · 2020
Later among the works it cites.
Iterative energy-based projection on a normal data manifold for anomaly localization
Dehaene, D., Frigo, O., Combrexelle, S., and Eline, P · 2020
Later among the works it cites.
Hyperplane arrangements of trained convnets are biased
Gamba, M., Carlsson, S., Azizpour, H., and Björkman, M · 2020
Later among the works it cites.
Grigsby, J. E. and Lindsey, K · 2020
Later among the works it cites.
Expression of fractals through neural network functions
Nadav, D., Barak, S., and Ingrid, D · 2020
Later among the works it cites.
Clustering earthquake signals and background noises in continuous seismic data with unsupervised deep learning
Seydoux, L., Balestriero, R., Poli, P., De Hoop, M., Campillo, M., and Baraniuk, R · 2020
Later among the works it cites.
Learning disconnected manifolds: a no gan’s land
Tanielian, U., Issenhuth, T., Dohmatob, E., and Mary, J · 2020
Later among the works it cites.
Orthogonal convolutional neural networks
Wang, J., Chen, Y., Chakraborty, R., and Yu, S. X · 2020
Later among the works it cites.