Skip to main navigation Skip to search Skip to main content

DYNAMIC MULTI-REFERENCE GENERATIVE PREDICTION FOR FACE VIDEO COMPRESSION

  • Zhao Wang*
  • , Bolin Chen*
  • , Yan Ye*
  • , Shiqi Wang
  • *Corresponding author for this work
  • Alibaba Group Holding Ltd.
  • City University of Hong Kong

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Face videos own abundant structured information and prior knowledge which can be utilized by generative neural networks to achieve ultra-low bitrate compression. However, generative neural network based face video compression suffers from large head motion which may easily result in deformed images. In this paper, the dynamic multi-reference prediction method is proposed for generative face video compression. Specifically, key map is extracted as the compact latent to represent the face image. The key maps of the current frame and multiple reference frames are used together to estimate multiple dense motion maps. The multiple motion maps are further applied to the corresponding reference frames to generate the final prediction of the current frame. Moreover, the reference frame can be dynamically refreshed during encoding to convert large head motion to relatively small motion. Experimental results show that the proposed method achieves superior compression performance compared to the state-of-the-art VVC standard as well as the latest generative face compression frameworks.

Original languageEnglish
Title of host publication2022 IEEE International Conference on Image Processing, ICIP 2022 - Proceedings
PublisherIEEE Computer Society
Pages896-900
Number of pages5
ISBN (Electronic)9781665496209
DOIs
StatePublished - 2022
Externally publishedYes
Event29th IEEE International Conference on Image Processing, ICIP 2022 - Bordeaux, France
Duration: 16 Oct 202219 Oct 2022

Publication series

NameProceedings - International Conference on Image Processing, ICIP
ISSN (Print)1522-4880

Conference

Conference29th IEEE International Conference on Image Processing, ICIP 2022
Country/TerritoryFrance
CityBordeaux
Period16/10/2219/10/22

Keywords

  • dynamic reference
  • Face video
  • generative network
  • multi reference
  • video compression

Fingerprint

Dive into the research topics of 'DYNAMIC MULTI-REFERENCE GENERATIVE PREDICTION FOR FACE VIDEO COMPRESSION'. Together they form a unique fingerprint.

Cite this