Share Email Print
cover

Proceedings Paper • new

Intra prediction with deep learning
Author(s): Raz Birman; Yoram Segal; Avishay David-Malka; Ofer Hadar
Format Member Price Non-Member Price
PDF $14.40 $18.00
cover GOOD NEWS! Your organization subscribes to the SPIE Digital Library. You may be able to download this paper for free. Check Access

Paper Abstract

One fundamental component of video compression standards is Intra-Prediction. Intra-Prediction takes advantage of redundancy in the information of neighboring pixel values within video frames to predict blocks of pixels from their surrounding pixels and thus allowing to transmit the prediction errors instead of the pixel values themselves. The prediction errors are of smaller values than the pixels themselves, thus allowing to accomplish compression of the video stream. Prevalent standards take advantage of intra-frame pixel value dependencies to perform prediction at the encoder end and transfer only residual errors to the decoder. The standards use multiple “Modes”, which are various linear combinations of pixels for prediction of their neighbors within image Macro-Blocks (MBs). In this research, we have used Deep Neural Networks (DNN) to perform the predictions. Using twelve Fully Connected Networks, we managed to reduce Mean Square Error (MSE) of the predicted error by up to 3 times as compared to standard modes prediction results. This substantial improvement comes at the expense of more extensive computations. However, these extra computations can be significantly mitigated by the use of dedicated Graphical Processing Units (GPUs).

Paper Details

Date Published: 17 September 2018
PDF: 9 pages
Proc. SPIE 10752, Applications of Digital Image Processing XLI, 1075214 (17 September 2018); doi: 10.1117/12.2320551
Show Author Affiliations
Raz Birman, Ben-Gurion Univ. of the Negev (Israel)
Yoram Segal, Ben-Gurion Univ. of the Negev (Israel)
Avishay David-Malka, Ben-Gurion Univ. of the Negev (Israel)
Ofer Hadar, Ben-Gurion Univ. of the Negev (Israel)


Published in SPIE Proceedings Vol. 10752:
Applications of Digital Image Processing XLI
Andrew G. Tescher, Editor(s)

© SPIE. Terms of Use
Back to Top