跳到主要导航 跳到搜索 跳到主要内容

Learning Spatio-Temporal Resolutions for Deep Video Compression

  • Jiancong Chen
  • , Meng Wang*
  • , Peilin Chen
  • , Shiqi Wang*
  • *此作品的通讯作者
  • City University of Hong Kong

科研成果: 期刊稿件文章同行评审

摘要

We propose a spatio-temporal adaptive deep video compression scheme, which is capable of intelligently adjusting the spatial resolution and temporal frame rate for content adaptive compression, with the aim of pursuing enhanced rate-distortion performance. In particular, a neural network-based spatio-temporal adaptation network is integrated into the deep video coding paradigm, enabling the adaptive determination of the optimal rescaling ratios for compression, leading to the further reduction of spatial and temporal redundancies. Moreover, learning-based modules for rescaling parameter determination are incorporated into the spatio-temporal adaptation network. The proposed scheme can be easily plugged into, and seamlessly collaborate with the existing deep video coding frameworks. Experimental results demonstrate that, compared to the original neural video codecs, the proposed method achieves significant bitrate savings in terms of both PSNR and MS-SSIM.

源语言英语
页(从-至)10493-10499
页数7
期刊IEEE Transactions on Circuits and Systems for Video Technology
35
10
DOI
出版状态已出版 - 2025
已对外发布

指纹

探究 'Learning Spatio-Temporal Resolutions for Deep Video Compression' 的科研主题。它们共同构成独一无二的指纹。

引用此