跳至主導覽 跳至搜尋 跳過主要內容

Controllable FrFT-Driven Y-Net for Cross-Modal Image Fusion Scenarios

  • Yifan Jiang
  • , Kemin Li
  • , Wenyuan Li
  • , Guoheng Huang
  • , Jietao Yang
  • , Xiaochen Yuan
  • , Xuhang Chen
  • , Bingo Wing Kuen Ling
  • , Chi Man Pun
  • , Shunlan Wang
  • Guangdong University of Technology
  • Huizhou University
  • Center for Integrated Circuits and Artificial Intelligence
  • University of Macau
  • Guangdong Provincial Hospital of Traditional Chinese Medicine

研究成果: Article同行評審

摘要

Image fusion aims to synthesize high-quality representations by integrating complementary information from multimodal sources. While recent deep learning approaches have explored spectral domain learning to overcome the localized receptive fields of spatial convolutions, they predominantly rely on the standard Fourier Transform. However, two critical bottlenecks remain: standard Fourier bases assume signal stationarity (ill-suited for non-stationary natural images) and lack adaptive frequency-adaptive feature weighting, leading to spectral leakage, ineffective high-frequency-noise decoupling, and texture smoothing. To break these limitations, we propose the Fractional-Order Dynamic Perceptive Y-Net (FDP Y-Net), a unified framework driven by Controllable Fractional Fourier Transform (FrFT) Convolutions. By generalizing the spectral domain, our model adaptively rotates the time-frequency axis to optimally represent non-stationary features and dynamically weights frequency components. Specifically, a Dynamic Perceptive Module spatially localizes salient regions, the Controllable FrFT Block captures time-varying spectral characteristics, and an Information Enhancement Module explicitly reconstructs high-frequency residuals. Orchestrated within a Y-shaped architecture with specialized skip connections, these components ensure holistic feature fusion. Extensive experiments across Infrared-Visible, Medical, and Multifocus datasets demonstrate that FDP Y-Net effectively resolves spectral limitations and frequency weighting deficiencies, achieving state-of-the-art performance in visual fidelity and quantitative metrics.

原文English
期刊IEEE Transactions on Consumer Electronics
DOIs
出版狀態Accepted/In press - 2026

指紋

深入研究「Controllable FrFT-Driven Y-Net for Cross-Modal Image Fusion Scenarios」主題。共同形成了獨特的指紋。

引用此