Fetching the paper…
Reading the bibliography…
Recurrent neural networks (RNNs), temporal convolutions, and neural differential equations (NDEs) are popular families of deep learning models for time-series data, each with unique strengths and tradeoffs in modeling power and computational efficiency.
Inverting modified matrices
Max A Woodbury · 1950
Earlier work this paper cites.
Orthogonal Polynomials
G. Szegö · 1967
Earlier work this paper cites.
Orthogonal Polynomials
G. Szegő · 1975
Earlier work this paper cites.
Approximation of dynamical systems by continuous time recurrent neural networks
Ken-ichi Funahashi and Yuichi Nakamura · 1993
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
On a new class of structured matrices
Y. Eidelman and I. Gohberg · 1999
Earlier work this paper cites.
Dynamic causal modelling
Karl J Friston, Lee Harrison, and Will Penny · 2003
Earlier work this paper cites.
Linear state-space control systems
Robert L Williams, Douglas A Lawrence, et al · 2007
Earlier work this paper cites.
Performance recovery in digital implementation of analogue systems
Guofeng Zhang, Tongwen Chen, and Xiang Chen · 2007
Earlier work this paper cites.
Numerical methods for ordinary differential equations , volume 2
John Charles Butcher and Nicolette Goodwin · 2008
Earlier work this paper cites.
A computational introduction to number theory and algebra
Victor Shoup · 2009
Earlier work this paper cites.
An introduction to orthogonal polynomials
T. S. Chihara · 2011
Earlier work this paper cites.
Orthogonal Polynomials and Related Approximation Results , pages 47–140
Jie Shen, Tao Tang, and Li-Lian Wang · 2011
Earlier work this paper cites.
The matrix cookbook, version 20121115
KB Petersen and MS Pedersen · 2012
Earlier work this paper cites.
Mathematical methods for physicists : a comprehensive guide / George B. Arfken, Hans J. Weber, Frank E. Harris
George B. (George Brown) Arfken, Hans Jürgen Weber, and Frank E Harris · 2013
Earlier work this paper cites.
Matrix computations , volume 3
Gene H Golub and Charles F Van Loan · 2013
Earlier work this paper cites.
On the difficulty of training recurrent neural networks
Razvan Pascanu, Tomas Mikolov, and Yoshua Bengio · 2013
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
A comprehensive review of stability analysis of continuous-time recurrent neural networks
Huaguang Zhang, Zhanshan Wang, and Derong Liu · 2014
Earlier work this paper cites.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Cited alongside, same era.
Unitary evolution recurrent neural networks
Martin Arjovsky, Amar Shah, and Yoshua Bengio · 2016
Cited alongside, same era.
Quasi-recurrent neural networks
James Bradbury, Stephen Merity, Caiming Xiong, and Richard Socher · 2016
Cited alongside, same era.
Computing with quasiseparable matrices
Clément Pernet · 2016
Cited alongside, same era.
Dilated recurrent neural networks
Shiyu Chang, Yang Zhang, Wei Han, Mo Yu, Xiaoxiao Guo, Wei Tan, Xiaodong Cui, Michael Witbrock, Mark Hasegawa-Johnson, and Thomas S Huang · 2017
Cited alongside, same era.
Xception: Deep learning with depthwise separable convolutions
François Chollet · 2017
Cited alongside, same era.
Gru-ode-bayes: Continuous modeling of sporadically-observed time series
Edward De Brouwer, Jaak Simm, Adam Arany, and Yves Moreau · 2019
Later among the works it cites.
Recurrent neural networks in the eye of differential equations
Murphy Yuezhen Niu, Lior Horesh, and Isaac Chuang · 2019
Later among the works it cites.
The pytorch-kaldi speech recognition toolkit
M. Ravanelli, T. Parcollet, and Y. Bengio · 2019
Later among the works it cites.
Latent ordinary differential equations for irregularly-sampled time series
Yulia Rubanova, Tian Qi Chen, and David K Duvenaud · 2019
Later among the works it cites.
Legendre memory units: Continuous-time representation in recurrent neural networks
Aaron Voelker, Ivana Kajić, and Chris Eliasmith · 2019
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Simple recurrent units for highly parallelizable recurrence
Tao Lei, Yu Zhang, Sida I Wang, Hui Dai, and Yoav Artzi · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
An empirical evaluation of generic convolutional and recurrent networks for sequence modeling
Shaojie Bai, J Zico Kolter, and Vladlen Koltun · 2018
Cited alongside, same era.
Recurrent neural networks for multivariate time series with missing values
Zhengping Che, Sanjay Purushotham, Kyunghyun Cho, David Sontag, and Yan Liu · 2018
Cited alongside, same era.
Neural ordinary differential equations
Tian Qi Chen, Yulia Rubanova, Jesse Bettencourt, and David K Duvenaud · 2018
Cited alongside, same era.
A two-pronged progress in structured dense matrix vector multiplication
Christopher De Sa, Albert Gu, Rohan Puttagunta, Christopher Ré, and Atri Rudra · 2018
Cited alongside, same era.
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Later among the works it cites.
Hippo: Recurrent memory with optimal polynomial projections
Albert Gu, Tri Dao, Stefano Ermon, Atri Rudra, and Christopher Ré · 2020
Later among the works it cites.
Rnns incrementally evolving on an equilibrium manifold: A panacea for vanishing and exploding gradients?
Anil Kag, Ziming Zhang, and Venkatesh Saligrama · 2020
Later among the works it cites.
Neural controlled differential equations for irregular time series
Patrick Kidger, James Morrill, James Foster, and Terry Lyons · 2020
Later among the works it cites.
Learning long-term dependencies in irregularly-sampled time series
Mathias Lechner and Ramin Hasani · 2020
Later among the works it cites.
Understanding the difficulty of training transformers
Liyuan Liu, Xiaodong Liu, Jianfeng Gao, Weizhu Chen, and Jiawei Han · 2020
Later among the works it cites.
Weak supervision as an efficient approach for automated seizure detection in electroencephalography
Khaled Saab, Jared Dunnmon, Christopher Ré, Daniel Rubin, and Christopher Lee-Messer · 2020
Later among the works it cites.
Parallelizing legendre memory unit training
Narsimha Chilkuri and Chris Eliasmith · 2021
Closest in time.
Catformer: Designing stable transformers via sensitivity analysis
Jared Quincy Davis, Albert Gu, Tri Dao, Krzysztof Choromanski, Christopher Ré, Percy Liang, and Chelsea Finn · 2021
Closest in time.
Lipschitz recurrent neural networks
N Benjamin Erichson, Omri Azencot, Alejandro Queiruga, Liam Hodgkinson, and Michael W Mahoney · 2021
Closest in time.
Gated recurrent units viewed through the lens of continuous time dynamical systems
Ian D Jordan, Piotr Aleksander Sokół, and Il Memming Park · 2021
Closest in time.
Neural rough differential equations for long time series
James Morrill, Cristopher Salvi, Patrick Kidger, James Foster, and Terry Lyons · 2021
Closest in time.
Time series extrinsic regression
Chang Wei Tan, Christoph Bergmeir, Francois Petitjean, and Geoffrey I Webb · 2021
Closest in time.