Fetching the paper…
Reading the bibliography…
Building on the interpretation of a recurrent neural network (RNN) as a continuous-time neural differential equation, we show, under appropriate conditions, that the solution of a RNN can be viewed as a linear function of a specific feature set of the input sequence, known as the signature.
Integration of paths–a faithful representation of paths by non-commutative formal power series
K.-T. Chen · 1958
Earlier work this paper cites.
An Introduction to Combinatorial Analysis
J. Riordan · 1958
Earlier work this paper cites.
The problem of learning long-term dependencies in recurrent networks
Y. Bengio, P. Frasconi, and P. Simard · 1993
Earlier work this paper cites.
On the derivatives of the sigmoid
A. A. Minai and R. D. Williams · 1993
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Runge-Kutta neural network for identification of dynamical systems in high accuracy
Y.-J. Wang and C.-T. Lin · 1998
Earlier work this paper cites.
Rademacher and Gaussian complexities: Risk bounds and structural results
P. L. Bartlett and S. Mendelson · 2002
Earlier work this paper cites.
Learning with kernels: support vector machines, regularization, optimization, and beyond
B. Schölkopf and A. J. Smola · 2002
Earlier work this paper cites.
Differential Equations Driven by Rough Paths , volume 1908 of Lecture Notes in Mathematics
T. J. Lyons, M. J. Caruana, and T. Lévy · 2007
Earlier work this paper cites.
Euler estimates for rough differential equations
P. Friz and N. Victoir · 2008
Earlier work this paper cites.
Kernel methods for deep learning
Y. Cho and L. Saul · 2009
Earlier work this paper cites.
Neural rough differential equations for long time series
J. Morrill, C. Salvi, P. Kidger, J. Foster, and T. Lyons · 2009
Earlier work this paper cites.
Multidimensional Stochastic Processes as Rough Paths: Theory and Applications , volume 120 of Cambridge Studies in Advanced Mathematics
P. K. Friz and N. B. Victoir · 2010
Earlier work this paper cites.
Recurrent neural network based language model
T. Mikolov, M. Karafiát, L. Burget, J. Černockỳ, and S. Khudanpur · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
R. Collobert, J. Weston, L. Bottou, M. Karlen, K. Kavukcuoglu, and P. Kuksa · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath, et al · 2012
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
A. Graves, A.-r. Mohamed, and G. Hinton · 2013
Earlier work this paper cites.
Learning from the past, predicting the statistics for the future, learning an evolving system
D. Levin, T. Lyons, and H. Ni · 2013
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Rough paths, signatures and the modelling of functions on streams
T. Lyons · 2014
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
D. P. Kingma and J. Ba · 2015
Cited alongside, same era.
A primer on the signature method in machine learning
I. Chevyrev and A. Kormilitzin · 2016
Cited alongside, same era.
DeepWriterID: An end-to-end online text-independent writer identification system
W. Yang, L. Jin, and M. Liu · 2016
Cited alongside, same era.
Invariance and stability of deep convolutional representations
A. Bietti and J. Mairal · 2017
Cited alongside, same era.
The Sacred Infrastructure for Computational Research
Klaus Greff, Aaron Klein, Martin Chovanec, Frank Hutter, and Jürgen Schmidhuber · 2017
Cited alongside, same era.
Kernels for sequentially ordered data
F. J. Király and H. Oberhauser · 2019
Later among the works it cites.
POPQORN: Quantifying robustness of recurrent neural networks
C.-Y. Ko, Z. Lyu, L. Weng, L. Daniel, N. Wong, and D. Lin · 2019
Later among the works it cites.
Learning stochastic differential equations using RNN with log signature features
S. Liao, T. Lyons, W. Yang, and H. Ni · 2019
Later among the works it cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Later among the works it cites.
Latent ordinary differential equations for irregularly-sampled time series
Y. Rubanova, R. T. Q. Chen, and D. K. Duvenaud · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Yang, T. Lyons, H. Ni, C. Schmid, and L. Jin · 2017
Cited alongside, same era.
To understand deep learning we need to understand kernel learning
M. Belkin, S. Ma, and S. Mandal · 2018
Cited alongside, same era.
Neural ordinary differential equations
R. T. Q. Chen, Y. Rubanova, J. Bettencourt, and D. K. Duvenaud · 2018
Cited alongside, same era.
Black-box generation of adversarial text sequences to evade deep learning classifiers
J. Gao, J. Lanchantin, M. L. Soffa, and Y. Qi · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Cited alongside, same era.
Towards deep learning models resistant to adversarial attacks
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu · 2018
Cited alongside, same era.
Derivatives pricing using signature payoffs
I. Perez Arribas · 2018
Cited alongside, same era.
Understanding generalization in recurrent neural networks
Z. Tu, F. He, and D. Tao · 2019
Later among the works it cites.
A path signature approach for speech emotion recognition
B. Wang, M. Liakata, H. Ni, T. Lyons, A. J. Nevado-Holgado, and K. Saunders · 2019
Later among the works it cites.
Computing the untruncated signature kernel as the solution of a Goursat problem
T. Cass, T. Lyons, C. Salvi, and W. Yang · 2020
Later among the works it cites.
On generalization bounds of a family of recurrent neural networks
M. Chen, X. Li, and T. Zhao · 2020
Later among the works it cites.
Theoretical guarantees for learning conditional expectation using controlled ODE-RNN
C. Herrera, F. Krach, and J. Teichmann · 2020
Later among the works it cites.
Learning differential equations that are easy to solve
J. Kelly, J. Bettencourt, M. J. Johnson, and D. K. Duvenaud · 2020
Later among the works it cites.
Neural controlled differential equations for irregular time series
P. Kidger, J. Morrill, J. Foster, and T. Lyons · 2020
Later among the works it cites.
Algorithm 1004: The iisignature library: Efficient calculation of iterated-integral signatures and log signatures
J. F. Reizenstein and B. Graham · 2020
Later among the works it cites.
Bayesian learning from sequential data using Gaussian processes with signature covariances
C. Toth and H. Oberhauser · 2020
Later among the works it cites.
SciPy 1.0: Fundamental Algorithms for Scientific Computing in Python
P. Virtanen, R. Gommers, T. E. Oliphant, M. Haberland, T. Reddy, D. Cournapeau, E. Burovski, P. Peterson, W. Weckesser, J. Bright, S. J. van der Walt, M. Brett, J. Wilson, K. J. Millman, N. Mayorov, A. R. J. Nelson, E. Jones, R. Kern, E. Larson, C. J. Carey, İ. Polat, Y. Feng, E. W. Moore, J. VanderPlas, D. Laxalde, J. Perktold, R. Cimrman, I. Henriksen, E. A. Quintero, C. R. Harris, A. M. Archibald, A. H. Ribeiro, F. Pedregosa, P. van Mulbregt, and SciPy 1.0 Contributors · 2020
Later among the works it cites.
Lipschitz recurrent neural networks
N. B. Erichson, O. Azencot, A. Queiruga, L. Hodgkinson, and M. W. Mahoney · 2021
Closest in time.
Embedding and learning with signatures
A. Fermanian · 2021
Closest in time.
Signatory: Differentiable computations of the signature and logsignature transforms, on both CPU and GPU
P. Kidger and T. Lyons · 2021
Closest in time.
Understanding recurrent neural networks using nonequilibrium response theory
S. H. Lim · 2021
Closest in time.