跳到主要导航 跳到搜索 跳到主要内容

Compact Temporal Trajectory Representation for Talking Face Video Compression

  • Bolin Chen
  • , Zhao Wang
  • , Binzhe Li
  • , Shiqi Wang*
  • , Yan Ye
  • *此作品的通讯作者
  • City University of Hong Kong
  • Peking University
  • Alibaba Group Holding Ltd.

科研成果: 期刊稿件文章同行评审

摘要

In this paper, we propose to compactly represent the nonlinear dynamics along the temporal trajectories for talking face video compression. By projecting the frames into a high dimensional space, the temporal trajectories of talking face frames, which are complex, non-linear and difficult to extrapolate, are implicitly modelled in an end-to-end inference framework based upon very compact feature representation. As such, the proposed framework is suitable for ultra-low bandwidth video communication and can guarantee the quality of the reconstructed video in such applications. The proposed compression scheme is also robust against large head-pose motions, due to the delicately designed dynamic reference refresh and temporal stabilization mechanisms. Experimental results demonstrate that compared to the state-of-the-art video coding standard Versatile Video Coding (VVC) as well as the latest generative compression schemes, our proposed scheme is superior in terms of both objective and subjective quality at the same bitrate. The project page can be found at https://github.com/Berlin0610/CTTR.

源语言英语
页(从-至)7009-7023
页数15
期刊IEEE Transactions on Circuits and Systems for Video Technology
33
11
DOI
出版状态已出版 - 1 11月 2023
已对外发布

指纹

探究 'Compact Temporal Trajectory Representation for Talking Face Video Compression' 的科研主题。它们共同构成独一无二的指纹。

引用此