Skip to main navigation Skip to search Skip to main content

Enhanced Motion-Compensated Video Coding With Deep Virtual Reference Frame Generation

  • Lei Zhao
  • , Shiqi Wang
  • , Xinfeng Zhang
  • , Shanshe Wang*
  • , Siwei Ma
  • , Wen Gao
  • *Corresponding author for this work

Research output: Journal Publications and ReviewsRGC 21 - Publication in refereed journalpeer-review

Abstract

In this paper, we propose an efficient inter prediction scheme by introducing the deep virtual reference frame (VRF), which serves better reference in the temporal redundancy removal process of video coding. In particular, the high quality VRF is generated with the deep learning-based frame rate up conversion (FRUC) algorithm from two reconstructed bi-directional frames, which is subsequently incorporated into the reference list serving as the high quality reference. Moreover, to alleviate the compression artifacts of VRF, we develop a convolutional neural network (CNN)-based enhancement model to further improve its quality. To facilitate better utilization of the VRF, a CTU level coding mode termed as direct virtual reference frame (DVRF) is devised, which achieves better trade-off between compression performance and complexity. The proposed scheme is integrated into HM-16.6 and JEM-7.1 software platforms, and the simulation results under random access (RA) configuration demonstrate significant superiority of the proposed method. When adding VRF to RPS, more than 6% average BD-rate gain is achieved for HEVC test sequences on HM-16.6, and 0.8% BD-rate gain is observed based on JEM-7.1 software. Regarding the DVRF mode, 3.6% bitrate saving is achieved on HM-16.6 with the computational complexity effectively reduced.
Original languageEnglish
Article number8704997
Pages (from-to)4832-4844
JournalIEEE Transactions on Image Processing
Volume28
Issue number10
Online published2 May 2019
DOIs
Publication statusPublished - Oct 2019

Research Keywords

  • deep learning
  • Inter prediction
  • video coding
  • virtual reference frame

Fingerprint

Dive into the research topics of 'Enhanced Motion-Compensated Video Coding With Deep Virtual Reference Frame Generation'. Together they form a unique fingerprint.

Cite this