Skip to main navigation Skip to search Skip to main content

Dual-Level Cross-Modality Neural Architecture Search for Guided Image Super-Resolution

  • Zhiwei Zhong
  • , Xianming Liu*
  • , Junjun Jiang
  • , Debin Zhao
  • , Shiqi Wang
  • *Corresponding author for this work
  • Faculty of Computing, Harbin Institute of Technology
  • City University of Hong Kong
  • Peng Cheng Laboratory

Research output: Contribution to journalArticlepeer-review

Abstract

Guided image super-resolution (GISR) aims to reconstruct a high-resolution (HR) target image from its low-resolution (LR) counterpart with the guidance of a HR image from another modality. Existing learning-based methods typically employ symmetric two-stream networks to extract features from both the guidance and target images, and then fuse these features at either an early or late stage through manually designed modules to facilitate joint inference. Despite significant performance, these methods still face several issues: i) the symmetric architectures treat images from different modalities equally, which may overlook the inherent differences between them; ii) lower-level features contain detailed information while higher-level features capture semantic structures. However, determining which layers should be fused and which fusion operations should be selected remain unresolved; iii) most methods achieve performance gains at the cost of increased computational complexity, so balancing the trade-off between computational complexity and model performance remains a critical issue. To address these issues, we propose a Dual-level Cross-modality Neural Architecture Search (DCNAS) framework to automatically design efficient GISR models. Specifically, we propose a dual-level search space that enables the NAS algorithm to identify effective architectures and optimal fusion strategies. Moreover, we propose a supernet training strategy that employs a pairwise ranking loss trained performance predictor to guide the supernet training process. To the best of our knowledge, this is the first attempt to introduce the NAS algorithm into GISR tasks. Extensive experiments demonstrate that the discovered model family, DCNAS-Tiny and DCNAS, achieve significant improvements on several GISR tasks, including guided depth map super-resolution, guided saliency map super-resolution, guided thermal image super-resolution, and pan-sharpening. Furthermore, we analyze the architectures searched by our method and provide some new insights for future research.

Original languageEnglish
Pages (from-to)8249-8267
Number of pages19
JournalIEEE Transactions on Pattern Analysis and Machine Intelligence
Volume47
Issue number9
DOIs
StatePublished - 2025
Externally publishedYes

Keywords

  • Guided image super-resolution
  • neural architecture search
  • ranking loss

Fingerprint

Dive into the research topics of 'Dual-Level Cross-Modality Neural Architecture Search for Guided Image Super-Resolution'. Together they form a unique fingerprint.

Cite this