Fetching the paper…
Reading the bibliography…
We show how bidirectional transformers trained for masked token prediction can be applied to neural image compression to achieve state-of-the-art results.
A learning algorithm for continually running fully recurrent neural networks
Ronald J Williams and David Zipser · 1989
Earlier work this paper cites.
Calculation of average psnr differences between rd-curves
Gisle Bjontegaard · 2001
Earlier work this paper cites.
Uniform distribution of sequences
Lauwerens Kuipers and Harald Niederreiter · 2012
Earlier work this paper cites.
End-to-end optimization of nonlinear transform codes for perceptual quality
Johannes Ballé, Valero Laparra, and Eero P Simoncelli · 2016
Earlier work this paper cites.
End-to-end optimized image compression
Johannes Ballé, Valero Laparra, and Eero P Simoncelli · 2016
Earlier work this paper cites.
Tim Salimans, Andrej Karpathy, Xi Chen, and Diederik P Kingma · 2017
Earlier work this paper cites.
Lossy image compression with compressive autoencoders
Lucas Theis, Wenzhe Shi, Andrew Cunningham, and Ferenc Huszar · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Variational image compression with a scale hyperprior
Johannes Ballé, David Minnen, Saurabh Singh, Sung Jin Hwang, and Nick Johnston · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Conditional probability models for deep image compression
Fabian Mentzer, Eirikur Agustsson, Michael Tschannen, Radu Timofte, and Luc Van Gool · 2018
Earlier work this paper cites.
Joint autoregressive and hierarchical priors for learned image compression
David Minnen, Johannes Ballé, and George D Toderici · 2018
Earlier work this paper cites.
Deep generative models for distribution-preserving lossy compression
Michael Tschannen, Eirikur Agustsson, and Mario Lucic · 2018
Earlier work this paper cites.
Generative adversarial networks for extreme learned image compression
Eirikur Agustsson, Michael Tschannen, Fabian Mentzer, Radu Timofte, and Luc Van Gool · 2019
Earlier work this paper cites.
Rethinking lossy compression: The rate-distortion-perception tradeoff
Yochai Blau and Tomer Michaeli · 2019
Cited alongside, same era.
Nonlinear transform coding
Johannes Ballé, Philip A Chou, David Minnen, Saurabh Singh, Nick Johnston, Eirikur Agustsson, Sung Jin Hwang, and George Toderici · 2020
Cited alongside, same era.
Learned image compression with discretized gaussian mixture likelihoods and attention modules
Zhengxue Cheng, Heming Sun, Masaru Takeuchi, and Jiro Katto · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Flax: A neural network library and ecosystem for JAX, 2020
Jonathan Heek, Anselm Levskaya, Avital Oliver, Marvin Ritter, Bertrand Rondepierre, Andreas Steiner, and Marc van Zee · 2020
Image compression with product quantized masked image modeling
Alaaeldin El-Nouby, Matthew J Muckley, Karen Ullrich, Ivan Laptev, Jakob Verbeek, and Hervé Jégou · 2022
Later among the works it cites.
VTM 17.1
Fraunhofer Gesellschaft · 2022
Later among the works it cites.
Elic: Efficient learned image compression with unevenly grouped space-channel contextual adaptive coding
Dailan He, Ziming Yang, Weikun Peng, Rui Ma, Hongwei Qin, and Yan Wang · 2022
Later among the works it cites.
http://r0k.us/graphics/kodak/ , 2022
Kodak PhotoCD dataset · 2022
Later among the works it cites.
A Burakhan Koyuncu, Han Gao, and Eckehard Steinbach · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
High-fidelity generative image compression
Fabian Mentzer, George D Toderici, Michael Tschannen, and Eirikur Agustsson · 2020
Cited alongside, same era.
Channel-wise autoregressive entropy models for learned image compression
David Minnen and Saurabh Singh · 2020
Cited alongside, same era.
On layer normalization in the transformer architecture
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tieyan Liu · 2020
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
Patrick Esser, Robin Rombach, and Bjorn Ommer · 2021
Cited alongside, same era.
Checkerboard context model for efficient learned image compression
Dailan He, Yaoyan Zheng, Baocheng Sun, Yan Wang, and Hongwei Qin · 2021
Cited alongside, same era.
Transformer-based transform coding
Yinhao Zhu, Yang Yang, and Taco Cohen · 2021
Cited alongside, same era.
Multi-realism image compression with a conditional generator
Eirikur Agustsson, David Minnen, George Toderici, and Fabian Mentzer · 2022
Cited alongside, same era.
Fabian Mentzer, George Toderici, David Minnen, Sung-Jin Hwang, Sergi Caelles, Mario Lucic, and Eirikur Agustsson · 2022
Later among the works it cites.
Entroformer: A transformer-based entropy model for learned image compression
Yichen Qian, Ming Lin, Xiuyu Sun, Zhiyu Tan, and Rong Jin · 2022
Later among the works it cites.
Phenaki: Variable length video generation from open domain textual description
Ruben Villegas, Mohammad Babaeizadeh, Pieter-Jan Kindermans, Hernan Moraldo, Han Zhang, Mohammad Taghi Saffar, Santiago Castro, Julius Kunze, and Dumitru Erhan · 2022
Later among the works it cites.
Mimt: Masked image modeling transformer for video compression
Jinxi Xiang, Kuan Tian, and Jun Zhang · 2022
Later among the works it cites.
An introduction to neural data compression
Y. Yang, S. Mandt, and L. Theis · 2022
Later among the works it cites.
The devil is in the details: Window-based attention for image compression
Renjie Zou, Chunfeng Song, and Zhaoxiang Zhang · 2022
Later among the works it cites.
Muse: Text-to-image generation via masked generative transformers
Huiwen Chang, Han Zhang, Jarred Barber, AJ Maschinot, Jose Lezama, Lu Jiang, Ming-Hsuan Yang, Kevin Murphy, William T Freeman, Michael Rubinstein, et al · 2023
Closest in time.
Unreasonable Effectiveness of Quasirandom Sequences
Martin Roberts · 2023
Closest in time.