科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Neural networks : the official journal of the International Neural Network Society2026-08-11

Sparse mixture of experts-driven multimodal degraded image fusion.

Yibing Yin, Nuo Chen, Sixiang Li, Yaoyue Gao, Yongjun Li

原始摘要(英文原文)· Original abstract
Multi-modal image fusion aims to integrate complementary information from different modal images to improve image quality and performance in visual tasks. However, existing methods face challenges such as the loss of key feature information in degraded scenarios, difficulty in adapting to multi-task fusion, and excessive computational overhead. To address these issues, this paper proposes an efficient fusion method. First, a staged collaborative optimization architecture is designed to decouple encoder pre-training from fusion layer fine-tuning, thereby enhancing single-modal feature representation and cross-modal semantic alignment. Second, a multi-scale heterogen eous hybrid expert architecture is proposed, integrated with a sparse activation mechanism, which dynamically selects the most relevant Top-K experts, significantly reducing redundant computations. Finally, a dynamic fusion paradigm selection mechanism is constructed, which adaptively selects the optimal fusion path based on input feature differences. We conducted extensive qualitative and quantitative experiments on the visible and infrared image fusion (VIF) Dataset and the Harvard Medical Dataset, validating the superior performance of this method on both datasets. Our method consistently outperforms existing SOTA methods in the core dimensions of 8 objective metrics on 6 key test datasets, especially showing superior performance in infrared-visible scenarios with low light, haze and noise interference, as well as noise-degraded medical imaging scenarios.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Sparse mixture of experts-driven multimodal degraded image fusion. — 科研速览 Science Skim