An End-to-End Compression Framework Based on Convolutional Neural Networks

被引:146
作者
Jiang, Feng [1 ]
Tao, Wen [1 ]
Liu, Shaohui [1 ]
Ren, Jie [1 ]
Guo, Xun [2 ]
Zhao, Debin [1 ]
机构
[1] Harbin Inst Technol, Sch Comp Sci & Technol, Harbin 150001, Heilongjiang, Peoples R China
[2] Microsoft Res Asia, Beijing 100080, Peoples R China
基金
中国国家自然科学基金;
关键词
Deep learning; compression framework; compact representation; convolutional neural networks (CNNs); ARTIFACT REDUCTION; DEBLOCKING; FILTER; DCT;
D O I
10.1109/TCSVT.2017.2734838
中图分类号
TM [电工技术]; TN [电子技术、通信技术];
学科分类号
0808 ; 0809 ;
摘要
Deep learning, e.g., convolutional neural networks (CNNs), has achieved great success in image processing and computer vision especially in high-level vision applications, such as recognition and understanding. However, it is rarely used to solve low-level vision problems such as image compression studied in this paper. Here, we move forward a step and propose a novel compression framework based on CNNs. To achieve high-quality image compression at low bit rates, two CNNs are seamlessly integrated into an end-to-end compression framework. The first CNN, named compact convolutional neural network (ComCNN), learns an optimal compact representation from an input image, which preserves the structural information and is then encoded using an image codes (e.g., JPEG, JPEG2000, or BPG). The second CNN, named reconstruction convolutional neural network (RecCNN), is used to reconstruct the decoded image with high quality in the decoding end. To make two CNNs effectively collaborate, we develop a unified end-to-end learning algorithm to simultaneously learn ComCNN and RecCNN, which facilitates the accurate reconstruction of the decoded image using RecCNN. Such a design also makes the proposed compression framework compatible with existing image coding standards. Experimental results validate that the proposed compression framework greatly outperforms several compression frameworks that use existing image coding standards with the state-of-the-art deblocking or denoising post-processing methods.
引用
收藏
页码:3007 / 3018
页数:12
相关论文
共 42 条
  • [1] [Anonymous], 2015, Accurate Image Super-Resolution Using Very Deep Convolutional Networks
  • [2] [Anonymous], 2017, LEARNING CONVOLUTION
  • [3] [Anonymous], 2012, ADADELTA ADAPTIVE LE
  • [4] [Anonymous], 2016, Full resolution image compression with recurrent neural networks
  • [5] [Anonymous], 2005, Live image quality assessment database release 2, DOI DOI 10.1109/CVPR.2015.7298594
  • [6] [Anonymous], 2016, PIXEL RECURRENT NEUR
  • [7] [Anonymous], 2015, Variable rate image compression with recurrent neural networks
  • [8] [Anonymous], 2015, IEEE I CONF COMP VIS, DOI DOI 10.1109/ICCV.2015.123
  • [9] [Anonymous], 2003, Standard codecs: Image compression to advanced video coding
  • [10] [Anonymous], 2017, LOSSY IMAGE COMPRESS