Fetching the paper…
Reading the bibliography…
The past few years have witnessed increasing interests in applying deep learning to video compression.
G. G. Langdon, “An introduction to arithmetic coding,” IBM Journal of Research and Development , vol. 28, no. 2, pp. 135–149, 1984
1984
Earlier work this paper cites.
N. Balakrishnan, Handbook of the logistic distribution . CRC Press, 1991
1991
Earlier work this paper cites.
D. J. Le Gall, “The MPEG video compression algorithm,” Signal Processing: Image Communication , vol. 4, no. 2, pp. 129–140, 1992
1992
Earlier work this paper cites.
G. E. Hinton and R. S. Zemel, “Autoencoders, minimum description length and helmholtz free energy,” in Advances in neural information processing systems , 1994, pp. 3–10
1994
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
A. Skodras, C. Christopoulos, and T. Ebrahimi, “The JPEG 2000 still image compression standard,” IEEE Signal processing magazine , vol. 18, no. 5, pp. 36–58, 2001
2001
Earlier work this paper cites.
G. Bjontegaard, “Calculation of average PSNR differences between RD-curves,” VCEG-M33 , 2001
2001
Earlier work this paper cites.
T. Wiegand, G. J. Sullivan, G. Bjontegaard, and A. Luthra, “Overview of the H.264/AVC video coding standard,” IEEE Transactions on circuits and systems for video technology , vol. 13, no. 7, pp. 560–576, 2003
2003
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Černockỳ, and S. Khudanpur, “Recurrent neural network based language model,” in Eleventh annual conference of the international speech communication association , 2010
2010
Earlier work this paper cites.
G. J. Sullivan, J.-R. Ohm, W.-J. Han, and T. Wiegand, “Overview of the high efficiency video coding (HEVC) standard,” IEEE Transactions on circuits and systems for video technology , vol. 22, no. 12, pp. 1649–1668, 2012
2012
Earlier work this paper cites.
K. Cho, “Simple sparsification improves sparse denoising autoencoders in denoising highly corrupted images,” in Proceedings of the International Conference on Machine Learning (ICML) , 2013, pp. 432–440
2013
Earlier work this paper cites.
F. Bossen, “Common test conditions and software reference configurations,” JCTVC-L1100 , vol. 12, 2013
2013
Earlier work this paper cites.
K. Cho, B. van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio, “Learning phrase representations using rnn encoder–decoder for statistical machine translation,” in Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2014, pp. 1724–1734
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in Advances in neural information processing systems , 2014, pp. 3104–3112
2014
Earlier work this paper cites.
K. Zeng, J. Yu, R. Wang, C. Li, and D. Tao, “Coupled deep autoencoder for single image super-resolution,” IEEE transactions on cybernetics , vol. 47, no. 1, pp. 27–37, 2015
2015
Earlier work this paper cites.
J. Donahue, L. Anne Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell, “Long-term recurrent convolutional networks for visual recognition and description,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 2625–2634
2015
Earlier work this paper cites.
O. Vinyals, A. Toshev, S. Bengio, and D. Erhan, “Show and tell: A neural image caption generator,” in Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR) , 2015, pp. 3156–3164
2015
Earlier work this paper cites.
N. Srivastava, E. Mansimov, and R. Salakhudinov, “Unsupervised learning of video representations using LSTMs,” in Proceedings of the International Conference on Machine Learning (ICML) , 2015, pp. 843–852
2015
Earlier work this paper cites.
S. Xingjian, Z. Chen, H. Wang, D.-Y. Yeung, W.-K. Wong, and W.-c. Woo, “Convolutional lstm network: A machine learning approach for precipitation nowcasting,” in Advances in neural information processing systems , 2015, pp. 802–810
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2015
2015
Earlier work this paper cites.
L. Gondara, “Medical image denoising using convolutional denoising autoencoders,” in 2016 IEEE 16th International Conference on Data Mining Workshops (ICDMW) . IEEE, 2016, pp. 241–246
2016
Earlier work this paper cites.
R. Wang and D. Tao, “Non-local auto-encoder with collaborative stabilization for image restoration,” IEEE Transactions on Image Processing , vol. 25, no. 5, pp. 2117–2129, 2016
2016
Cited alongside, same era.
A. Karpathy, J. Johnson, and L. Fei-Fei, “Visualizing and understanding recurrent networks,” 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
G. Toderici, S. M. O’Malley, S. J. Hwang, D. Vincent, D. Minnen, S. Baluja, M. Covell, and R. Sukthankar, “Variable rate image compression with recurrent neural networks,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2016
2016
Cited alongside, same era.
N. Johnston, D. Vincent, D. Minnen, M. Covell, S. Singh, T. Chinen, S. Jin Hwang, J. Shor, and G. Toderici, “Improved lossy image compression with priming and spatially adaptive bit rates for recurrent networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 4385–4393
2018
Later among the works it cites.
M. Xu, T. Li, Z. Wang, X. Deng, R. Yang, and Z. Guan, “Reducing complexity of HEVC: A deep learning approach,” IEEE Transactions on Image Processing , vol. 27, no. 10, pp. 5044–5059, 2018
2018
Later among the works it cites.
J. Liu, S. Xia, W. Yang, M. Li, and D. Liu, “One-for-all: Grouped variation network-based fractional interpolation in video coding,” IEEE Transactions on Image Processing , vol. 28, no. 5, pp. 2140–2151, 2018
2018
Later among the works it cites.
J. Lee, S. Cho, and S.-K. Beack, “Context-adaptive entropy model for end-to-end optimized image compression,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Van den Oord, N. Kalchbrenner, L. Espeholt, O. Vinyals, A. Graves et al. , “Conditional image generation with pixelcnn decoders,” in Advances in neural information processing systems , 2016, pp. 4790–4798
2016
Cited alongside, same era.
H. Wang, W. Gan, S. Hu, J. Y. Lin, L. Jin, L. Song, P. Wang, I. Katsavounidis, A. Aaron, and C.-C. J. Kuo, “MCL-JCV: a JND-based H.264/AVC video quality assessment dataset,” in Proceedings of the IEEE International Conference on Image Processing (ICIP) . IEEE, 2016, pp. 1509–1513
2016
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in Neural Information Processing Systems (NeurIPS) , 2017, pp. 5998–6008
2016
Cited alongside, same era.
K. G. Lore, A. Akintayo, and S. Sarkar, “Llnet: A deep autoencoder approach to natural low-light image enhancement,” Pattern Recognition , vol. 61, pp. 650–662, 2017
2017
Cited alongside, same era.
G. Toderici, D. Vincent, N. Johnston, S. Jin Hwang, D. Minnen, J. Shor, and M. Covell, “Full resolution image compression with recurrent neural networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 5306–5314
2017
Cited alongside, same era.
E. Agustsson, F. Mentzer, M. Tschannen, L. Cavigelli, R. Timofte, L. Benini, and L. V. Gool, “Soft-to-hard vector quantization for end-to-end learning compressible representations,” in Advances in Neural Information Processing Systems (NeurIPS) , 2017, pp. 1141–1151
2017
Cited alongside, same era.
L. Theis, W. Shi, A. Cunningham, and F. Huszár, “Lossy image compression with compressive autoencoders,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
J. Ballé, V. Laparra, and E. P. Simoncelli, “End-to-end optimized image compression,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2017
2017
Cited alongside, same era.
2019
Later among the works it cites.
G. Lu, W. Ouyang, D. Xu, X. Zhang, C. Cai, and Z. Gao, “DVC: An end-to-end deep video compression framework,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 11 006–11 015
2019
Later among the works it cites.
Z. Cheng, H. Sun, M. Takeuchi, and J. Katto, “Learning image and video compression through spatial-temporal energy compaction,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 10 071–10 080
2019
Later among the works it cites.
A. Djelouah, J. Campos, S. Schaub-Meyer, and C. Schroers, “Neural inter-frame compression for video coding,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 6421–6429
2019
Later among the works it cites.
H. Choi and I. V. Bajić, “Deep frame prediction for video coding,” IEEE Transactions on Circuits and Systems for Video Technology , 2019
2019
Later among the works it cites.
T. Li, M. Xu, R. Yang, and X. Tao, “A DenseNet based approach for multi-frame in-loop filter in HEVC,” in Proceedings of the Data Compression Conference (DCC) . IEEE, 2019, pp. 270–279
2019
Later among the works it cites.
T. Li, M. Xu, C. Zhu, R. Yang, Z. Wang, and Z. Guan, “A deep learning approach for multi-frame in-loop filter of HEVC,” IEEE Transactions on Image Processing , 2019
2019
Later among the works it cites.
Z. Chen, T. He, X. Jin, and F. Wu, “Learning for video compression,” IEEE Transactions on Circuits and Systems for Video Technology , 2019
2019
Later among the works it cites.
A. Habibian, T. van Rozendaal, J. M. Tomczak, and T. S. Cohen, “Video compression with rate-distortion autoencoders,” in Proceedings of the IEEE International Conference of Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
T. Xue, B. Chen, J. Wu, D. Wei, and W. T. Freeman, “Video enhancement with task-oriented flow,” International Journal of Computer Vision , vol. 127, no. 8, pp. 1106–1125, 2019
2019
Later among the works it cites.
Y. Hu, W. Yang, and J. Liu, “Coarse-to-fine hyper-prior modeling for learned image compression,” in Proceedings of the AAAI Conference on Artificial Intelligence , 2020
2020
Closest in time.
R. Yang, F. Mentzer, L. Van Gool, and R. Timofte, “Learning for video compression with hierarchical quality and recurrent enhancement,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Closest in time.
H. Liu, L. Huang, M. Lu, T. Chen, and Z. Ma, “Learned video compression via joint spatial-temporal correlation exploration,” in Proceedings of the AAAI Conference on Artificial Intelligence , 2020
2020
Closest in time.
E. Agustsson, D. Minnen, N. Johnston, J. Balle, S. J. Hwang, and G. Toderici, “Scale-space flow for end-to-end optimized video compression,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2020, pp. 8503–8512
2020
Closest in time.
2020
Closest in time.
A. Mercat, M. Viitanen, and J. Vanne, “UVG dataset: 50/120fps 4K sequences for video codec analysis and development,” in Proceedings of the 11th ACM Multimedia Systems Conference , 2020, pp. 297–302
2020
Closest in time.
Cisco, “Cisco visual networking index: Global mobile data traffic forecast update, 2017-2022 white paper,” https://www.cisco.com/c/en/us/solutions/collateral/service-provider/visual-networking-index-vni/white-paper-c11-738429.html
2022
Closest in time.