跳到主要导航 跳到搜索 跳到主要内容

VDMamba: Vector Decomposition in Vision Mamba for Image Deraining and Beyond

  • Kui Jiang
  • , Junjun Jiang
  • , Shiqi Wang
  • , Wenqi Ren
  • , Chia Wen Lin
  • , Zhengguo Li
  • Harbin Institute of Technology
  • Open Research Fund from Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)
  • School of Computer Science and Technology, Harbin Institute of Technology
  • City University of Hong Kong
  • Sun Yat-Sen University
  • National Tsing Hua University
  • Agency for Science, Technology and Research, Singapore

科研成果: 期刊稿件文章同行评审

摘要

Recent research employs Mamba for image restoration, yet the intrinsic coupling between deraining characteristics and Mamba architectures remains underexplored. We propose VDMamba, a vector decomposition-based vision Mamba approach that leverages 1D sequential representations to characterize direction-aware rain distributions in the frequency embedding space. The core innovation is the Mamba-based Vector Decomposition and Synthesis Module (VDSM). VDSM derives vertical and horizontal 1D vectors from frequency components and utilizes single-direction Mamba scanning to eliminate direction-specific perturbations. This enables the exploration of global relationships for accurate learning without complex scanning designs. Additionally, these components are encoded via bidirectional coupling for refinement. Experiments on various tasks, including deraining, dehazing, and low-light enhancement, demonstrate VDMamba’s competitive performance. Specifically, it achieves a 0.58 dB PSNR improvement in deraining compared to the NeRD method, while reducing model parameters by 94.3%, computational cost by 88.3%, and inference time by 77.5%.

源语言英语
页(从-至)3339-3352
页数14
期刊IEEE Transactions on Multimedia
28
DOI
出版状态已出版 - 2026
已对外发布

指纹

探究 'VDMamba: Vector Decomposition in Vision Mamba for Image Deraining and Beyond' 的科研主题。它们共同构成独一无二的指纹。

引用此