Fetching the paper…
Reading the bibliography…
Semantically meaningful information content in perceptual signals is usually unevenly distributed.
A thermionic trigger
Schmitt, O. H · 1938
Earlier work this paper cites.
Results of a prototype television bandwidth compression scheme
Robinson, A. H. and Cherry, C · 1967
Earlier work this paper cites.
Neural sequence chunkers
Schmidhuber, J · 1991
Earlier work this paper cites.
Acceleration of stochastic approximation by averaging
Polyak, B. T. and Juditsky, A. B · 1992
Earlier work this paper cites.
Nonlinear total variation based noise removal algorithms
Rudin, L. I., Osher, S., and Fatemi, E · 1992
Earlier work this paper cites.
Quantization
Gray, R. M. and Neuhoff, D. L · 1998
Earlier work this paper cites.
Slow feature analysis: Unsupervised learning of invariances
Wiskott, L. and Sejnowski, T. J · 2002
Earlier work this paper cites.
Infinite latent feature models and the indian buffet process
Ghahramani, Z. and Griffiths, T · 2005
Earlier work this paper cites.
Regularization and variable selection via the elastic net
Zou, H. and Hastie, T · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A., Fernández, S., Gomez, F., and Schmidhuber, J · 2006
Earlier work this paper cites.
Matplotlib: A 2d graphics environment
Hunter, J. D · 2007
Earlier work this paper cites.
A maximum-likelihood interpretation for slow feature analysis
Turner, R. and Sahani, M · 2007
Earlier work this paper cites.
Learning invariant features through topographic filter maps
Kavukcuoglu, K., Ranzato, M., Fergus, R., and LeCun, Y · 2009
Earlier work this paper cites.
Python 3 Reference Manual
Van Rossum, G. and Drake, F. L · 2009
Earlier work this paper cites.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Léonard, N., and Courville, A · 2013
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
Earlier work this paper cites.
Learning ordered representations with nested dropout
Rippel, O., Gelbart, M., and Adams, R · 2014
Earlier work this paper cites.
SoX - Sound eXchange, 2015
Bagwell, C. et al · 2015
Earlier work this paper cites.
Scheduled sampling for sequence prediction with recurrent neural networks
Bengio, S., Vinyals, O., Jaitly, N., and Shazeer, N · 2015
Earlier work this paper cites.
The unreasonable effectiveness of recurrent neural networks, 2015
Karpathy, A · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Earlier work this paper cites.
librosa: Audio and music signal analysis in python
McFee, B., Raffel, C., Liang, D., Ellis, D. P., McVicar, M., Battenberg, E., and Nieto, O · 2015
Earlier work this paper cites.
Sequence level training with recurrent neural networks
Ranzato, M., Chopra, S., Auli, M., and Zaremba, W · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Abadi, M., Barham, P., Chen, J., Chen, Z., Davis, A., Dean, J., Devin, M., Ghemawat, S., Irving, G., Isard, M., et al · 2016
Earlier work this paper cites.
Generating sentences from a continuous space
Bowman, S. R., Vilnis, L., Vinyals, O., Dai, A. M., Józefowicz, R., and Bengio, S · 2016
Earlier work this paper cites.
An infinite restricted boltzmann machine
Côté, M.-A. and Larochelle, H · 2016
Earlier work this paper cites.
Attend, infer, repeat: Fast scene understanding with generative models
Eslami, S. A., Heess, N., Weber, T., Tassa, Y., Szepesvari, D., Kavukcuoglu, K., and Hinton, G. E · 2016
Earlier work this paper cites.
Adaptive computation time for recurrent neural networks
Graves, A · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
End-to-end learning of action detection from frame glimpses in videos
Yeung, S., Russakovsky, O., Mori, G., and Fei-Fei, L · 2016
Earlier work this paper cites.
End-to-end optimized image compression
Ballé, J., Laparra, V., and Simoncelli, E. P · 2017
Earlier work this paper cites.
Skip rnn: Learning to skip state updates in recurrent neural networks
Campos, V., Jou, B., i Nieto, X. G., Torres, J., and Chang, S.-F · 2017
Earlier work this paper cites.
Variational lossy autoencoder
Chen, X., Kingma, D. P., Salimans, T., Duan, Y., Dhariwal, P., Schulman, J., Sutskever, I., and Abbeel, P · 2017
Earlier work this paper cites.
Spatially adaptive computation time for residual networks
Figurnov, M., Collins, M. D., Zhu, Y., Zhang, L., Huang, J., Vetrov, D., and Salakhutdinov, R · 2017
Cited alongside, same era.
Group sparse optimization via lp,q regularization
Hu, Y., Li, C., Meng, K., Qin, J., and Yang, X · 2017
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Jang, E., Gu, S., and Poole, B · 2017
Cited alongside, same era.
Sgdr: Stochastic gradient descent with warm restarts
Loshchilov, I. and Hutter, F · 2017
Cited alongside, same era.
The concrete distribution: A continuous relaxation of discrete random variables
Maddison, C. J., Mnih, A., and Teh, Y. W · 2017
Cited alongside, same era.
Pointer sentinel mixture models
Merity, S., Xiong, C., Bradbury, J., and Socher, R · 2017
Cited alongside, same era.
Invariant information clustering for unsupervised image classification and segmentation
Ji, X., Vedaldi, A., and Henriques, J · 2019
Later among the works it cites.
Efficient segmentation: Learning downsampling near semantic boundaries
Marin, D., He, Z., Vajda, P., Chatterjee, P., Tsai, S., Yang, F., and Boykov, Y · 2019
Later among the works it cites.
Musenet, 2019
Payne, C · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., and Sutskever, I · 2019
Later among the works it cites.
Latent ordinary differential equations for irregularly-sampled time series
Rubanova, Y., Chen, R. T. Q., and Duvenaud, D. K · 2019
Later among the works it cites.
Interpolation-prediction networks for irregularly sampled time series
Shukla, S. N. and Marlin, B · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Blind phoneme segmentation with temporal prediction errors
Michel, P., Rasanen, O., Thiollière, R., and Dupoux, E · 2017
Cited alongside, same era.
Discrete event, continuous time rnns
Mozer, M. C., Kazakov, D., and Lindsey, R. V · 2017
Cited alongside, same era.
Stick-breaking variational autoencoders
Nalisnick, E. and Smyth, P · 2017
Cited alongside, same era.
Phased lstm: Accelerating recurrent network training for long or event-based sequences
Neil, D., Pfeiffer, M., and Liu, S.-C · 2017
Cited alongside, same era.
Discrete variational autoencoders
Rolfe, J. T · 2017
Cited alongside, same era.
Performance rnn: Generating music with expressive timing and dynamics
Simon, I. and Oore, S · 2017
Cited alongside, same era.
Videobert: A joint model for video and language representation learning
Sun, C., Myers, A., Vondrick, C., Murphy, K., and Schmid, C · 2019
Later among the works it cites.
The bitter lesson, 2019
Sutton, R · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z., Dai, Z., Yang, Y., Carbonell, J., Salakhutdinov, R. R., and Le, Q. V · 2019
Later among the works it cites.
Libritts: A corpus derived from librispeech for text-to-speech
Zen, H., Clark, R., Weiss, R. J., Dang, V., Jia, Y., Wu, Y., Zhang, Y., and Chen, Z · 2019
Later among the works it cites.
The DeepMind JAX Ecosystem, 2020
Babuschkin, I., Baumli, K., Bell, A., Bhupatiraju, S., Bruce, J., Buchlovsky, P., Budden, D., Cai, T., Clark, A., Danihelka, I., Fantacci, C., Godwin, J., Jones, C., Hennigan, T., Hessel, M., Kapturowski, S., Keck, T., Kemaev, I., King, M., Martens, L., Mikulik, V., Norman, T., Quan, J., Papamakarios, G., Ring, R., Ruiz, F., Sanchez, A., Schneider, R., Sezener, E., Spencer, S., Srinivasan, S., Stokowiec, W., and Viola, F · 2020
Later among the works it cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., Agarwal, S., Herbert-Voss, A., Krueger, G., Henighan, T., Child, R., Ramesh, A., Ziegler, D. M., Wu, J., Winter, C., Hesse, C., Chen, M., Sigler, E., Litwin, M., Gray, S., Chess, B., Clark, J., Berner, C., McCandlish, S., Radford, A., Sutskever, I., and Amodei, D · 2020
Later among the works it cites.
Jukebox: A generative model for music
Dhariwal, P., Jun, H., Payne, C., Kim, J. W., Radford, A., and Sutskever, I · 2020
Later among the works it cites.
Ethnologue: Languages of the World. Twenty-third edition
Eberhard, D. M., Simons, G. F., and Fennig, C. D · 2020
Later among the works it cites.
Event-based vision: A survey
Gallego, G., Delbruck, T., Orchard, G. M., Bartolozzi, C., Taba, B., Censi, A., Leutenegger, S., Davison, A., Conradt, J., Daniilidis, K., et al · 2020
Later among the works it cites.
Vector quantized contrastive predictive coding for template-based music generation
Hadjeres, G. and Crestel, L · 2020
Later among the works it cites.
Array programming with NumPy
Harris, C. R., Millman, K. J., van der Walt, S. J., Gommers, R., Virtanen, P., Cournapeau, D., Wieser, E., Taylor, J., Berg, S., Smith, N. J., Kern, R., Picus, M., Hoyer, S., van Kerkwijk, M. H., Brett, M., Haldane, A., del R’ıo, J. F., Wiebe, M., Peterson, P., G’erard-Marchant, P., Sheppard, K., Reddy, T., Weckesser, W., Abbasi, H., Gohlke, C., and Oliphant, T. E · 2020
Later among the works it cites.
Discretalk: Text-to-speech as a machine translation problem
Hayashi, T. and Watanabe, S · 2020
Later among the works it cites.
Libri-light: A benchmark for asr with limited or no supervision
Kahn, J., Rivière, M., Zheng, W., Kharitonov, E., Xu, Q., Mazaré, P.-E., Karadayi, J., Liptchinsky, V., Collobert, R., Fuegen, C., et al · 2020
Later among the works it cites.
A convolutional deep markov model for unsupervised speech representation learning
Khurana, S., Laurent, A., Hsu, W.-N., Chorowski, J., Łańcucki, A., Marxer, R., and Glass, J · 2020
Later among the works it cites.
Pointrend: Image segmentation as rendering
Kirillov, A., Wu, Y., He, K., and Girshick, R · 2020
Later among the works it cites.
Object-centric learning with slot attention
Locatello, F., Weissenborn, D., Unterthiner, T., Mahendran, A., Heigold, G., Uszkoreit, J., Dosovitskiy, A., and Kipf, T · 2020
Later among the works it cites.
Surprisal-triggered conditional computation with neural networks
Lugosch, L., Nowrouzezahrai, D., and Meyer, B. H · 2020
Later among the works it cites.
This time with feeling: Learning expressive musical performance
Oore, S., Simon, I., Dieleman, S., Eck, D., and Simonyan, K · 2020
Later among the works it cites.
Stabilizing transformers for reinforcement learning
Parisotto, E., Song, F., Rae, J., Pascanu, R., Gulcehre, C., Jayakumar, S., Jaderberg, M., Kaufman, R. L., Clark, A., Noury, S., et al · 2020
Later among the works it cites.
Zero: Memory optimizations toward training trillion parameter models
Rajbhandari, S., Rasley, J., Ruwase, O., and He, Y · 2020
Later among the works it cites.
Scipy 1.0: Fundamental algorithms for scientific computing in python
Virtanen, P., Gommers, R., Oliphant, T. E., Haberland, M., Reddy, T., Cournapeau, D., Burovski, E., Peterson, P., Weckesser, W., Bright, J., van der Walt, S. J., Brett, M., Wilson, J., Jarrod Millman, K., Mayorov, N., Nelson, A. R. J., Jones, E., Kern, R., Larson, E., Carey, C., Polat, l., Feng, Y., Moore, E. W., Vand erPlas, J., Laxalde, D., Perktold, J., Cimrman, R., Henriksen, I., Quintero, E. A., Harris, C. R., Archibald, A. M., Ribeiro, A. H., Pedregosa, F., van Mulbregt, P., and Contributors, S. . · 2020
Later among the works it cites.
Visual transformers: Token-based image representation and processing for computer vision
Wu, B., Xu, C., Dai, X., Wan, A., Zhang, P., Tomizuka, M., Keutzer, K., and Vajda, P · 2020
Later among the works it cites.
Uwspeech: Speech to speech translation for unwritten languages
Zhang, C., Tan, X., Ren, Y., Qin, T., Zhang, K., and Liu, T.-Y · 2020
Later among the works it cites.
Generative speech coding with predictive variance regularization
Kleijn, W. B., Storus, A., Chinen, M., Denton, T., Lim, F. S., Luebs, A., Skoglund, J., and Yeh, H · 2021
Closest in time.
Towards nonlinear disentanglement in natural data with temporal sparse coding
Klindt, D. A., Schott, L., Sharma, Y., Ustyuzhaninov, I., Brendel, W., Bethge, M., and Paiton, D · 2021
Closest in time.
Generative spoken language modeling from raw audio
Lakhotia, K., Kharitonov, E., Hsu, W.-N., Adi, Y., Polyak, A., Bolte, B., Nguyen, T.-A., Copet, J., Baevski, A., Mohamed, A., et al · 2021
Closest in time.
Generating images with sparse representations
Nash, C., Menick, J., Dieleman, S., and Battaglia, P · 2021
Closest in time.
Zero-shot text-to-image generation
Ramesh, A., Pavlov, M., Goh, G., Gray, S., Voss, C., Radford, A., Chen, M., and Sutskever, I · 2021
Closest in time.
Differentiable segmentation of sequences
Scharwächter, E., Lennartz, J., and Müller, E · 2021
Closest in time.
Unsupervised speech representation learning using wavenet autoencoders
Chorowski, J., Weiss, R. J., Bengio, S., and van den Oord, A · 2053
Closest in time.