Fetching the paper…
Reading the bibliography…
Neural networks inspired by differential equations have proliferated for the past several years.
L. Pontryagin, E. Mishchenko, V. Boltyanski, and R. Gamkrelidze, The mathematical theory of optimal processes . Interscience Publishers, 1962
1962
Earlier work this paper cites.
J. Dormand and P. Prince, “A family of embedded runge-kutta formulae,” Journal of Computational and Applied Mathematics , vol. 6, no. 1, pp. 19 – 26, 1980
1980
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, pp. 1735–80, 12 1997
1997
Earlier work this paper cites.
T. J. Lyons, “Differential equations driven by rough signals.” Revista Matemática Iberoamericana , vol. 14, no. 2, pp. 215–310, 1998
1998
Earlier work this paper cites.
M. Giles and N. Pierce, “An introduction to the adjoint approach to design,” Flow, Turbulence and Combustion , vol. 65, pp. 393–415, 2000
2000
Earlier work this paper cites.
W. Hager, “Runge-kutta methods in optimal control and the transformed adjoint system,” Numerische Mathematik , vol. 87, pp. 247–282, 2000
2000
Earlier work this paper cites.
G. C. Reinsel, Elements of multivariate time series analysis . Springer Science & Business Media, 2003
2003
Earlier work this paper cites.
M. W. Spratling and M. H. Johnson, “A feedback model of visual attention,” Journal of cognitive neuroscience , vol. 16, no. 2, pp. 219–237, 2004
2004
Earlier work this paper cites.
T. Lyons, M. Caruana, and T. Lévy, Differential Equations Driven by Rough Paths . Springer, 2004, École D’Eté de Probabilités de Saint-Flour XXXIV - 2004
2004
Earlier work this paper cites.
C. A. Ralanamahatana, J. Lin, D. Gunopulos, E. Keogh, M. Vlachos, and G. Das, “Mining time series data,” in Data mining and knowledge discovery handbook . Springer, 2005, pp. 1069–1103
2005
Earlier work this paper cites.
P. J. Reiter, “Using cart to generate partially synthetic, public use microdata,” Journal of Official Statistics , vol. 21, p. 441, 01 2005
2005
Earlier work this paper cites.
N. K. Ahmed, A. F. Atiya, N. E. Gayar, and H. El-Shishiny, “An empirical comparison of machine learning models for time series forecasting,” Econometric Reviews , vol. 29, no. 5-6, pp. 594–621, 2010
2010
Earlier work this paper cites.
B. Krollner, B. J. Vanstone, and G. R. Finnie, “Financial time series forecasting with machine learning techniques: a survey.” in ESANN , 2010
2010
Earlier work this paper cites.
T.-c. Fu, “A review on time series data mining,” Engineering Applications of Artificial Intelligence , vol. 24, no. 1, pp. 164–181, 2011
2011
Earlier work this paper cites.
P. Esling and C. Agon, “Time-series data mining,” ACM Computing Surveys (CSUR) , vol. 45, no. 1, pp. 1–34, 2012
2012
Earlier work this paper cites.
G. Kirchgässner, J. Wolters, and U. Hassler, Introduction to modern time series analysis . Springer Science & Business Media, 2012
2012
Earlier work this paper cites.
G. Bontempi, S. B. Taieb, and Y.-A. Le Borgne, “Machine learning strategies for time series forecasting,” in European business intelligence summer school . Springer, 2012, pp. 62–77
2012
Earlier work this paper cites.
2014
Cited alongside, same era.
D. Bahdanau, K. Cho, and Y. Bengio, “Neural machine translation by jointly learning to align and translate,” in ICLR , 2015
2015
Cited alongside, same era.
K. Cho, A. Courville, and Y. Bengio, “Describing multimedia content using attention-based encoder-decoder networks,” IEEE Transactions on Multimedia , 2015
2015
Cited alongside, same era.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in ICML , 2015
2015
Cited alongside, same era.
Q. You, H. Jin, Z. Wang, C. Fang, and J. Luo, “Image captioning with semantic attention,” in CVPR , 2016
2019
Later among the works it cites.
H. Kim, A. Mnih, J. Schwarz, M. Garnelo, A. Eslami, D. Rosenbaum, O. Vinyals, and Y. W. Teh, “Attentive neural processes,” in ICLR , 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
J. B. Lee, R. A. Rossi, S. Kim, N. K. Ahmed, and E. Koh, “Attention models in graphs: A survey,” ACM Trans. Knowl. Discov. Data , vol. 13, no. 6, 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
J. Lu, J. Yang, D. Batra, and D. Parikh, “Hierarchical question-image co-attention for visual question answering,” in NeurIPS , 2016
2016
Cited alongside, same era.
Z. Che, S. Purushotham, K. Cho, D. Sontag, and Y. Liu, “Recurrent neural networks for multivariate time series with missing values,” 2016
2016
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, u. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017
2017
Cited alongside, same era.
A. S. Weigend, Time series prediction: forecasting the future and understanding the past . Routledge, 2018
2018
Cited alongside, same era.
Z. Che, S. Purushotham, K. Cho, D. Sontag, and Y. Liu, “Recurrent neural networks for multivariate time series with missing values,” Scientific reports , vol. 8, no. 1, pp. 1–12, 2018
2018
Cited alongside, same era.
R. T. Q. Chen, Y. Rubanova, J. Bettencourt, and D. K. Duvenaud, “Neural ordinary differential equations,” in NeurIPS , 2018
2018
Cited alongside, same era.
T. Shen, T. Zhou, G. Long, J. Jiang, S. Pan, and C. Zhang, “Disan: Directional self-attention network for rnn/cnn-free language understanding,” in AAAI , 2018
2018
Cited alongside, same era.
M. A. Reyna, C. Josef, S. Seyedi, R. Jeter, S. P. Shashikumar, M. Brandon Westover, A. Sharma, S. Nemati, and G. D. Clifford, “Early prediction of sepsis from clinical data: the physionet/computing in cardiology challenge 2019,” in 2019 Computing in Cardiology (CinC) , 2019, pp. Page 1–Page 4
2019
Later among the works it cites.
J. Yoon, D. Jarrett, and M. van der Schaar, “Time-series generative adversarial networks,” in NeurIPS , 2019
2019
Later among the works it cites.
I. D. Jordan, P. A. Sokol, and I. M. Park, “Gated recurrent units viewed through the lens of continuous time dynamical systems,” 2019
2019
Later among the works it cites.
Y. Rubanova, R. T. Q. Chen, and D. Duvenaud, “Latent odes for irregularly-sampled time series,” 2019
2019
Later among the works it cites.
P. Kidger, J. Morrill, J. Foster, and T. J. Lyons, “Neural controlled differential equations for irregular time series,” in NeurIPS , 2020
2020
Later among the works it cites.
C. Zang and F. Wang, “Neural dynamics on complex networks,” in Proceedings of the 26th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , 2020, pp. 892–902
2020
Later among the works it cites.
P. Kidger, J. Morrill, J. Foster, and T. Lyons, “Neural controlled differential equations for irregular time series,” in NeurIPS , 2020
2020
Later among the works it cites.
J. Zhuang, N. Dvornek, X. Li, S. Tatikonda, X. Papademetris, and J. Duncan, “Adaptive checkpoint adjoint method for gradient estimation in neural ode,” in ICML , 2020
2020
Later among the works it cites.
A. Galassi, M. Lippi, and P. Torroni, “Attention in natural language processing,” IEEE Transactions on Neural Networks and Learning Systems , 2020
2020
Later among the works it cites.
P. Gao, X. Yang, R. Zhang, and K. Huang, “Explainable tensorized neural ordinary differential equations forarbitrary-step time series prediction,” 2020
2020
Later among the works it cites.
C. Herrera, F. Krach, and J. Teichmann, “Neural jump ordinary differential equations: Consistent continuous-time prediction and filtering,” in ICLR , 2021
2021
Closest in time.
S. Y. Jhin, M. Jo, T. Kong, J. Jeon, and N. Park, “Ace-node: Attentive co-evolving neural ordinary differential equations,” in KDD , 2021
2021
Closest in time.