Fetching the paper…
Reading the bibliography…
While recent machine learning research has revealed connections between deep generative models such as VAEs and rate-distortion losses used in learned compression, most of this work has focused on images.
C. C. Cutler, “Differential quantization of communication signals,” Jul. 29 1952, uS Patent 2,605,361
1952
Earlier work this paper cites.
R. E. Kalman et al. , “A new approach to linear filtering and prediction problems [j],” Journal of basic Engineering , vol. 82, no. 1, pp. 35–45, 1960
1960
Earlier work this paper cites.
G. J. Sullivan and T. Wiegand, “Rate-distortion optimization for video compression,” IEEE signal processing magazine , vol. 15, no. 6, pp. 74–90, 1998
1998
Earlier work this paper cites.
C. A. Glasbey and K. V. Mardia, “A review of image-warping methods,” Journal of applied statistics , vol. 25, no. 2, pp. 155–171, 1998
1998
Earlier work this paper cites.
T. M. Cover, Elements of information theory . John Wiley & Sons, 1999
1999
Earlier work this paper cites.
V. K. Goyal, “Theoretical foundations of transform coding,” IEEE Signal Processing Magazine , vol. 18, no. 5, pp. 9–21, 2001
2001
Earlier work this paper cites.
T. Wiegand, G. J. Sullivan, G. Bjontegaard, and A. Luthra, “Overview of the h. 264/avc video coding standard,” IEEE Transactions on circuits and systems for video technology , vol. 13, no. 7, pp. 560–576, 2003
2003
Earlier work this paper cites.
Z. Wang, E. P. Simoncelli, and A. C. Bovik, “Multiscale structural similarity for image quality assessment,” in The Thrity-Seventh Asilomar Conference on Signals, Systems & Computers, 2003 , vol. 2. Ieee, 2003, pp. 1398–1402
2003
Earlier work this paper cites.
G. J. Sullivan, J.-R. Ohm, W.-J. Han, and T. Wiegand, “Overview of the high efficiency video coding (hevc) standard,” IEEE Transactions on circuits and systems for video technology , vol. 22, no. 12, pp. 1649–1668, 2012
2012
Earlier work this paper cites.
A. Gersho and R. M. Gray, Vector quantization and signal compression . Springer Science & Business Media, 2012, vol. 159
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
B. Uria, I. Murray, and H. Larochelle, “Rnade: The real-valued neural autoregressive density-estimator,” Advances in Neural Information Processing Systems , vol. 26, 2013
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
F. Bellard, “Bpg image format,” 2014. [Online]. Available: https://bellard.org/bpg/bpg_spec.txt
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Chung, K. Kastner, L. Dinh, K. Goel, A. C. Courville, and Y. Bengio, “A recurrent latent variable model for sequential data,” in Advances in neural information processing systems , 2015, pp. 2980–2988
2015
Earlier work this paper cites.
K. Sohn, H. Lee, and X. Yan, “Learning structured output representation using deep conditional generative models,” in Advances in Neural Information Processing Systems , vol. 28, 2015
2015
Earlier work this paper cites.
A. Dosovitskiy, P. Fischer, E. Ilg, P. Hausser, C. Hazirbas, V. Golkov, P. Van Der Smagt, D. Cremers, and T. Brox, “Flownet: Learning optical flow with convolutional networks,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2758–2766
2015
Earlier work this paper cites.
D. J. Rezende and S. Mohamed, “Variational inference with normalizing flows,” in Proceedings of the 32nd International Conference on International Conference on Machine Learning-Volume 37 , 2015, pp. 1530–1538
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Ballé, V. Laparra, and E. P. Simoncelli, “End-to-end optimization of nonlinear transform codes for perceptual quality,” 2016
2016
Earlier work this paper cites.
C. Vondrick, H. Pirsiavash, and A. Torralba, “Generating videos with scene dynamics,” in Advances in neural information processing systems , 2016, pp. 613–621
2016
Earlier work this paper cites.
D. P. Kingma, T. Salimans, R. Jozefowicz, X. Chen, I. Sutskever, and M. Welling, “Improved variational inference with inverse autoregressive flow,” in Advances in neural information processing systems , 2016, pp. 4743–4751
2016
Earlier work this paper cites.
H. Wang, W. Gan, S. Hu, J. Y. Lin, L. Jin, L. Song, P. Wang, I. Katsavounidis, A. Aaron, and C.-C. J. Kuo, “Mcl-jcv: a jnd-based h. 264/avc video quality assessment dataset,” in 2016 IEEE International Conference on Image Processing (ICIP) . IEEE, 2016, pp. 1509–1513
2016
Earlier work this paper cites.
G. Papamakarios, T. Pavlakou, and I. Murray, “Masked autoregressive flow for density estimation,” in Advances in Neural Information Processing Systems , 2017, pp. 2338–2347
2017
Earlier work this paper cites.
J. Ballé, V. Laparra, and E. P. Simoncelli, “End-to-end optimized image compression,” in 5th International Conference on Learning Representations, ICLR 2017 , 2017
2017
Cited alongside, same era.
L. Theis, W. Shi, A. Cunningham, and F. Huszár, “Lossy image compression with compressive autoencoders,” International Conference on Learning Representations , 2017
2017
Cited alongside, same era.
G. Toderici, D. Vincent, N. Johnston, S. Jin Hwang, D. Minnen, J. Shor, and M. Covell, “Full resolution image compression with recurrent neural networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , July 2017
2017
Cited alongside, same era.
T. Chen, H. Liu, Q. Shen, T. Yue, X. Cao, and Z. Ma, “Deepcoder: A deep neural network based video compression,” in 2017 IEEE Visual Communications and Image Processing (VCIP) , 2017, pp. 1–4
2017
Cited alongside, same era.
Z. Chen, T. He, X. Jin, and F. Wu, “Learning for video compression,” IEEE Transactions on Circuits and Systems for Video Technology , vol. 30, no. 2, pp. 566–576, 2019
2019
Later among the works it cites.
T. Xue, B. Chen, J. Wu, D. Wei, and W. T. Freeman, “Video enhancement with task-oriented flow,” International Journal of Computer Vision (IJCV) , vol. 127, no. 8, pp. 1106–1125, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Cited alongside, same era.
D. Minnen, J. Ballé, and G. D. Toderici, “Joint autoregressive and hierarchical priors for learned image compression,” in Advances in Neural Information Processing Systems , 2018, pp. 10 771–10 780
2018
Cited alongside, same era.
D. P. Kingma and P. Dhariwal, “Glow: Generative flow with invertible 1x1 convolutions,” in Advances in Neural Information Processing Systems , 2018, pp. 10 215–10 224
2018
Cited alongside, same era.
F. Schmidt and T. Hofmann, “Deep state space models for unconditional word generation,” in Advances in Neural Information Processing Systems , 2018, pp. 6158–6168
2018
Cited alongside, same era.
E. Denton and R. Fergus, “Stochastic video generation with a learned prior,” in International Conference on Machine Learning . PMLR, 2018, pp. 1174–1183
2018
Cited alongside, same era.
Y. Li and S. Mandt, “Disentangled sequential autoencoder,” in Proceedings of the 35th International Conference on Machine Learning . PMLR, 2018, pp. 5670–5679
2018
Cited alongside, same era.
N. Johnston, D. Vincent, D. Minnen, M. Covell, S. Singh, T. Chinen, S. Jin Hwang, J. Shor, and G. Toderici, “Improved lossy image compression with priming and spatially adaptive bit rates for recurrent networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 4385–4393
2018
Cited alongside, same era.
J. Ballé, D. Minnen, S. Singh, S. J. Hwang, and N. Johnston, “Variational image compression with a scale hyperprior,” International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
2020
Later among the works it cites.
E. Agustsson, D. Minnen, N. Johnston, J. Balle, S. J. Hwang, and G. Toderici, “Scale-space flow for end-to-end optimized video compression,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 8503–8512
2020
Later among the works it cites.
J. Marino, L. Chen, J. He, and S. Mandt, “Improving sequential latent variable models with autoregressive flows,” in Symposium on Advances in Approximate Bayesian Inference , 2020, pp. 1–16
2020
Later among the works it cites.
H. Liu, H. Shen, L. Huang, M. Lu, T. Chen, and Z. Ma, “Learned video compression via joint spatial-temporal correlation exploration,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, 2020, pp. 11 580–11 587
2020
Later among the works it cites.
J. Lin, D. Liu, H. Li, and F. Wu, “M-lvc: Multiple frames prediction for learned video compression,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 3546–3554
2020
Later among the works it cites.
G. Lu, C. Cai, X. Zhang, L. Chen, W. Ouyang, D. Xu, and Z. Gao, “Content adaptive and error propagation aware deep video compression,” in European Conference on Computer Vision . Springer, 2020, pp. 456–472
2020
Later among the works it cites.
R. Yang, Y. Yang, J. Marino, Y. Yang, and S. Mandt, “Deep generative video compression with temporal autoregressive transforms,” ICML 2020 Workshop on Invertible Neural Networks, Normalizing Flows, and Explicit Likelihood Models , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Behrmann, P. Vicol, K.-C. Wang, R. Grosse, and J.-H. Jacobsen, “Understanding and mitigating exploding inverses in invertible neural networks,” 2020
2020
Later among the works it cites.
D. Minnen and S. Singh, “Channel-wise autoregressive entropy models for learned image compression,” in 2020 IEEE International Conference on Image Processing (ICIP) . IEEE, 2020, pp. 3339–3343
2020
Later among the works it cites.
Y. Yang, R. Bamler, and S. Mandt, “Variational bayesian quantization,” in International Conference on Machine Learning , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
R. Yang, F. Mentzer, L. Van Gool, and R. Timofte, “Learning for video compression with hierarchical quality and recurrent enhancement,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Later among the works it cites.
A. Mercat, M. Viitanen, and J. Vanne, “Uvg dataset: 50/120fps 4k sequences for video codec analysis and development,” in Proceedings of the 11th ACM Multimedia Systems Conference , 2020, pp. 297–302
2020
Later among the works it cites.
Y. Yang, G. Sautière, J. J. Ryu, and T. S. Cohen, “Feedback recurrent autoencoder,” in ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2020, pp. 3347–3351
2020
Later among the works it cites.
G. Papamakarios, E. Nalisnick, D. J. Rezende, S. Mohamed, and B. Lakshminarayanan, “Normalizing flows for probabilistic modeling and inference,” 2021
2021
Closest in time.
R. Yang, F. Mentzer, L. Van Gool, and R. Timofte, “Learning for video compression with recurrent auto-encoder and recurrent probability model,” IEEE Journal of Selected Topics in Signal Processing , vol. 15, no. 2, pp. 388–401, 2021
2021
Closest in time.
2022
Closest in time.
2022
Closest in time.
C.-W. Huang, D. Krueger, A. Lacoste, and A. Courville, “Neural autoregressive flows,” in International Conference on Machine Learning . PMLR, 2018, pp. 2078–2087
2087
Closest in time.