Fetching the paper…
Reading the bibliography…
In this paper, we tackle the important yet under-investigated problem of making long-horizon prediction of event sequences.
Spectra of some self-exciting and mutually exciting point processes
Hawkes, A. G · 1971
Earlier work this paper cites.
Simulation of nonhomogeneous Poisson processes by thinning
Lewis, P. A. and Shedler, G. S · 1979
Earlier work this paper cites.
Neural networks and physical systems with emergent collective computationalabilities
Hopfield, J · 1982
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Training products of experts by minimizing contrastive divergence
Hinton, G. E · 2002
Earlier work this paper cites.
A tutorial on energy-based learning
LeCun, Y., Chopra, S., Hadsell, R., Ranzato, M., and Huang, F.-J · 2006
Earlier work this paper cites.
An Introduction to the Theory of Point Processes, Volume II: General Theory and Structure
Daley, D. J. and Vere-Jones, D · 2007
Earlier work this paper cites.
A unified energy-based framework for unsupervised learning
Ranzato, M., Boureau, Y.-L., Chopra, S., and LeCun, Y · 2007
Earlier work this paper cites.
Multivariate Hawkes processes
Liniger, T. J · 2009
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Gutmann, M. and Hyvärinen, A · 2010
Earlier work this paper cites.
Learning deep energy models
Ngiam, J., Chen, Z., Koh, P. W., and Ng, A. Y · 2011
Earlier work this paper cites.
A fast and simple algorithm for training neural probabilistic language models
Mnih, A. and Teh, Y. W · 2012
Earlier work this paper cites.
Training energy-based models for time-series imputation
Brakel, P., Stroobandt, D., and Schrauwen, B · 2013
Earlier work this paper cites.
SNAP Datasets: Stanford large network dataset collection , 2014
Leskovec, J. and Krevl, A · 2014
Earlier work this paper cites.
FOILing NYC’s taxi trip data , 2014
Whong, C · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. and Ba, J · 2015
Earlier work this paper cites.
Recurrent marked temporal point processes: Embedding event history to vector
Du, N., Dai, H., Trivedi, R., Upadhyay, U., Gomez-Rodriguez, M., and Song, L · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Oord, A. v. d., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., and Kavukcuoglu, K · 2016
Cited alongside, same era.
Sequence level training with recurrent neural networks
Ranzato, M., Chopra, S., Auli, M., and Zaremba, W · 2016
Cited alongside, same era.
The neural Hawkes process: A neurally self-modulating multivariate point process
Mei, H. and Eisner, J · 2017
Cited alongside, same era.
Automatic differentiation in PyTorch
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A · 2017
Cited alongside, same era.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I · 2017
User-dependent neural sequence models for continuous-time event data
Boyd, A., Bamler, R., Mandt, S., and Smyth, P · 2020
Later among the works it cites.
Residual energy-based models for text generation
Deng, Y., Bakhtin, A., Ott, M., Szlam, A., and Ranzato, M · 2020
Later among the works it cites.
Neural temporal point processes [for] modelling electronic health records
Enguehard, J., Busbridge, D., Bozson, A., Woodcock, C., and Hammerla, N · 2020
Later among the works it cites.
Intensity-free learning of temporal point processes
Shchur, O., Biloš, M., and Günnemann, S · 2020
Later among the works it cites.
Self-attentive Hawkes process
Zhang, Q., Lipani, A., Kirnap, O., and Yilmaz, E · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
User behavior data from taobao for recommendation , 2018
Alibaba · 2018
Cited alongside, same era.
Long text generation via adversarial training with leaked information
Guo, J., Lu, S., Cai, H., Zhang, W., Yu, Y., and Wang, J · 2018
Cited alongside, same era.
Ma, Z. and Collins, M · 2018
Cited alongside, same era.
Real or fake? learning to discriminate machine from human generated text
Bakhtin, A., Gross, S., Ott, M., Deng, Y., Ranzato, M., and Szlam, A · 2019
Cited alongside, same era.
Implicit generation and modeling with energy based models
Du, Y. and Mordatch, I · 2019
Cited alongside, same era.
Shape and time distortion loss for training deep time series forecasting models
Le Guen, V. and Thome, N · 2019
Cited alongside, same era.
Imputing missing events in continuous-time event streams
Mei, H., Qin, G., and Eisner, J · 2019
Cited alongside, same era.
Zuo, S., Jiang, H., Li, Z., Zhao, T., and Zha, H · 2020
Later among the works it cites.
Long horizon forecasting with temporal point processes
Deshpande, P., Marathe, K., De, A., and Sarawagi, S · 2021
Later among the works it cites.
Characterizing and Overcoming the Limitations of Neural Autoregressive Models
Goyal, K · 2021
Later among the works it cites.
Long text generation by modeling sentence-level and discourse-level coherence
Guan, J., Mao, X., Fan, C., Liu, Z., Ding, W., and Huang, M · 2021
Later among the works it cites.
Limitations of autoregressive models and their alternatives
Lin, C.-C., Jaech, A., Li, X., Gormley, M., and Eisner, J · 2021
Later among the works it cites.
What context features can transformer language models use?
O’Connor, J. and Andreas, J · 2021
Later among the works it cites.
Trajectory prediction with latent belief energy-based model
Pang, B., Zhao, T., Xie, X., and Wu, Y · 2021
Later among the works it cites.
Identifying coordinated accounts on social media through hidden influence and group behaviours
Sharma, K., Zhang, Y., Ferrara, E., and Liu, Y · 2021
Later among the works it cites.
Deep Fourier kernel for self-attentive point processes
Zhu, S., Zhang, M., Ding, R., and Xie, Y · 2021
Later among the works it cites.
On the uncomputability of partition functions in energy-based sequence models
Lin, C.-C. and McCarthy, A. D · 2022
Closest in time.
Transformer embeddings of irregularly spaced events and their participants
Yang, C., Mei, H., and Eisner, J · 2022
Closest in time.