跳到主要导航 跳到搜索 跳到主要内容

Rethinking Semantic Image Compression: Scalable Representation With Cross-Modality Transfer

  • Pingping Zhang
  • , Shiqi Wang*
  • , Meng Wang
  • , Jiguo Li
  • , Xu Wang
  • , Sam Kwong
  • *此作品的通讯作者
  • City University of Hong Kong
  • University of Chinese Academy of Sciences
  • Fudan University
  • Shenzhen University

科研成果: 期刊稿件文章同行评审

摘要

This article proposes the scalable cross-modality compression (SCMC) paradigm, in which the image compression problem is further cast into a representation task by hierarchically sketching the image with different modalities. Herein, we adopt the conceptual organization philosophy to model the overwhelmingly complicated visual patterns, based upon the semantic, structure, and signal level representation accounting for different tasks. The SCMC paradigm that incorporates the representation at different granularities supports diverse application scenarios, such as high-level semantic communication and low-level image reconstruction. The decoder, which enables the recovery of the visual information, benefits from the scalable coding based upon the semantic, structure, and signal layers. Qualitative and quantitative results demonstrate that the SCMC can convey accurate semantic and perceptual information of images, especially at low bitrates, and promising rate-distortion performance has been achieved compared to state-of-the-art methods. The code will be available online https://github.com/ppingzhang/SCMC.

源语言英语
页(从-至)4441-4445
页数5
期刊IEEE Transactions on Circuits and Systems for Video Technology
33
8
DOI
出版状态已出版 - 1 8月 2023
已对外发布

指纹

探究 'Rethinking Semantic Image Compression: Scalable Representation With Cross-Modality Transfer' 的科研主题。它们共同构成独一无二的指纹。

引用此