Extreme Generative Image Compression by Learning Text Embedding from Diffusion Models

11/14/2022
by   Zhihong Pan, et al.
0

Transferring large amount of high resolution images over limited bandwidth is an important but very challenging task. Compressing images using extremely low bitrates (<0.1 bpp) has been studied but it often results in low quality images of heavy artifacts due to the strong constraint in the number of bits available for the compressed data. It is often said that a picture is worth a thousand words but on the other hand, language is very powerful in capturing the essence of an image using short descriptions. With the recent success of diffusion models for text-to-image generation, we propose a generative image compression method that demonstrates the potential of saving an image as a short text embedding which in turn can be used to generate high-fidelity images which is equivalent to the original one perceptually. For a given image, its corresponding text embedding is learned using the same optimization process as the text-to-image diffusion model itself, using a learnable text embedding as input after bypassing the original transformer. The optimization is applied together with a learning compression model to achieve extreme compression of low bitrates <0.1 bpp. Based on our experiments measured by a comprehensive set of image quality metrics, our method outperforms the other state-of-the-art deep learning methods in terms of both perceptual quality and diversity.

READ FULL TEXT

page 2

page 3

page 4

page 6

page 7

research
05/26/2023

High-Fidelity Image Compression with Score-based Generative Models

Despite the tremendous success of diffusion generative models in text-to...
research
06/14/2020

CompressNet: Generative Compression at Extremely Low Bitrates

Compressing images at extremely low bitrates (< 0.1 bpp) has always been...
research
07/17/2023

Extreme Image Compression using Fine-tuned VQGAN Models

Recent advances in generative compression methods have demonstrated rema...
research
05/23/2022

Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

We present Imagen, a text-to-image diffusion model with an unprecedented...
research
01/13/2023

Neural Image Compression with a Diffusion-Based Decoder

Diffusion probabilistic models have recently achieved remarkable success...
research
05/04/2023

Multi-Modality Deep Network for JPEG Artifacts Reduction

In recent years, many convolutional neural network-based models are desi...
research
01/18/2018

Near-lossless L-infinity constrained Multi-rate Image Decompression via Deep Neural Network

Recently a number of CNN-based techniques were proposed to remove image ...

Please sign up or login with your details

Forgot password? Click here to reset