Pith. sign in

DDFM: Denoising Diffusion Model for Multi-Modality Image Fusion

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Multi-modality image fusion aims to combine different modalities to produce fused images that retain the complementary features of each modality, such as functional highlights and texture details. To leverage strong generative priors and address challenges such as unstable training and lack of interpretability for GAN-based generative methods, we propose a novel fusion algorithm based on the denoising diffusion probabilistic model (DDPM). The fusion task is formulated as a conditional generation problem under the DDPM sampling framework, which is further divided into an unconditional generation subproblem and a maximum likelihood subproblem. The latter is modeled in a hierarchical Bayesian manner with latent variables and inferred by the expectation-maximization (EM) algorithm. By integrating the inference solution into the diffusion sampling iteration, our method can generate high-quality fused images with natural image generative priors and cross-modality information from source images. Note that all we required is an unconditional pre-trained generative model, and no fine-tuning is needed. Our extensive experiments indicate that our approach yields promising fusion results in infrared-visible image fusion and medical image fusion. The code is available at \url{https://github.com/Zhaozixiang1228/MMIF-DDFM}.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Task-driven Image Fusion with Learnable Fusion Loss

cs.CV · 2024-12-04 · conditional · novelty 4.0

TDFusion learns the per-pixel weights of a fusion loss from the downstream task loss using MAML-style inner and outer updates, improving infrared-visible fusion for segmentation and detection.

citing papers explorer

Showing 1 of 1 citing paper.

  • Task-driven Image Fusion with Learnable Fusion Loss cs.CV · 2024-12-04 · conditional · none · ref 84 · internal anchor

    TDFusion learns the per-pixel weights of a fusion loss from the downstream task loss using MAML-style inner and outer updates, improving infrared-visible fusion for segmentation and detection.