Timezone: »
Recent models for learned image compression are based on autoencoders that learn approximately invertible mappings from pixels to a quantized latent representation. The transforms are combined with an entropy model, which is a prior on the latent representation that can be used with standard arithmetic coding algorithms to generate a compressed bitstream. Recently, hierarchical entropy models were introduced as a way to exploit more structure in the latents than previous fully factorized priors, improving compression performance while maintaining end-to-end optimization. Inspired by the success of autoregressive priors in probabilistic generative models, we examine autoregressive, hierarchical, and combined priors as alternatives, weighing their costs and benefits in the context of image compression. While it is well known that autoregressive models can incur a significant computational penalty, we find that in terms of compression performance, autoregressive and hierarchical priors are complementary and can be combined to exploit the probabilistic structure in the latents better than all previous learned models. The combined model yields state-of-the-art rate-distortion performance and generates smaller files than existing methods: 15.8% rate reductions over the baseline hierarchical model and 59.8%, 35%, and 8.4% savings over JPEG, JPEG2000, and BPG, respectively. To the best of our knowledge, our model is the first learning-based method to outperform the top standard image codec (BPG) on both the PSNR and MS-SSIM distortion metrics.
Author Information
David Minnen (Google)
Johannes Ballé (Google)
Johannes Ballé (Google)
George D Toderici (Google)
More from the Same Authors
-
2022 Poster: VCT: A Video Compression Transformer »
Fabian Mentzer · George D Toderici · David Minnen · Sergi Caelles · Sung Jin Hwang · Mario Lucic · Eirikur Agustsson -
2021 : Invited talk #10: Johannes Ballé »
Johannes Ballé -
2020 Poster: An Unsupervised Information-Theoretic Perceptual Quality Metric »
Sangnie Bhardwaj · Ian Fischer · Johannes Ballé · Troy Chinen -
2020 Poster: High-Fidelity Generative Image Compression »
Fabian Mentzer · George D Toderici · Michael Tschannen · Eirikur Agustsson -
2020 Oral: High-Fidelity Generative Image Compression »
Fabian Mentzer · George D Toderici · Michael Tschannen · Eirikur Agustsson -
2017 Poster: Eigen-Distortions of Hierarchical Representations »
Alexander Berardino · Valero Laparra · Johannes Ballé · Eero Simoncelli -
2017 Oral: Eigen-Distortions of Hierarchical Representations »
Alexander Berardino · Valero Laparra · Johannes Ballé · Eero Simoncelli