跳到主要导航 跳到搜索 跳到主要内容

Monotonic and Invertible Network: A General Framework for Learning IQA Model from Mixed Datasets

  • Baoliang Chen
  • , Kang Xiao
  • , Xuelin Shen*
  • , Shiqi Wang*
  • *此作品的通讯作者
  • South China Normal University
  • Lingnan University
  • Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)
  • City University of Hong Kong

科研成果: 期刊稿件文章同行评审

摘要

Learning from mixed datasets is a powerful strategy for model generalization improvement. However, this approach becomes particularly challenging for image quality assessment (IQA) where the quality annotations across different datasets are usually not aligned, due to varying quality criteria, score ranges, and viewing conditions involved in their subjective tests. Score rescaling and rank learning are two main strategies for the mixed dataset IQA model learning. In particular, the score rescaling attempts to align the annotations across datasets directly by empirical linear or nonlinear transformations, which may not always be reliable. Rank learning, on the other hand, is restricted to pair comparison within each dataset while the image pairs from different datasets are not fully examined. In this paper, we present a novel mixed dataset learning framework for the IQA, where we align the quality annotation in an implicit manner. Specifically, we first introduce a dataset-shared quality regressor to project images across datasets into a unified proxy score space. We force the proxy scores to serve as proxies of the aligned results of the annotations. To achieve this, two key priors are considered: 1) within each dataset, the proxy scores should maintain the same rank as the annotations; and 2) the score ranges of proxy scores from different datasets overlap when images of similar quality exist in these datasets. To meet the criteria, we propose a monotonic and invertible network as a dataset-specific score mapper to bridge the proxy scores and the annotations with the rank consistency and range intersection established. This constrained learning ultimately results in the proxy scores being the desired alignment results of the annotations. Experiments on extensive IQA datasets validate the effectiveness of our method, and the continuous performance gains when incorporating our learning strategy into different network architectures demonstrate the high generalizability of our framework. The source code is available at https://github.com/KANGX99/MIMI.

源语言英语
页(从-至)7924-7945
页数22
期刊International Journal of Computer Vision
133
11
DOI
出版状态已出版 - 11月 2025
已对外发布

指纹

探究 'Monotonic and Invertible Network: A General Framework for Learning IQA Model from Mixed Datasets' 的科研主题。它们共同构成独一无二的指纹。

引用此