科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Optics Express2026-01-12· Hyperspectral imaging

MMCTNet: multimodal cross-scale transformer network for hyperspectral and LiDAR/SAR image classification

Songpeng Gong, Uzair Aslam Bhatti, Yonis Gulzar, Mohammad Shuaib Mir, Harold Neira-Molina, Ayoob Lone

原始摘要(英文原文)· Original abstract
This paper addresses the challenge of multisource feature fusion for hyperspectral images (HSI) and light detection and ranging (LiDAR)/synthetic aperture radar (SAR) in complex scene classification tasks. A multimodal cross-scale transformer network (MMCTNet) is proposed to tackle this issue. The model employs a spatial self-attention (SSA) module to enhance intra-modal spatial dependency modeling, while a multiscale adaptive fusion (MSAF) module achieves cross-modal semantic complementation. Furthermore, a transformer encoder combined with a cross-attention mechanism facilitates global semantic interaction and feature collaboration. Experiments conducted on four public datasets-MUUFL, Augsburg, Berlin, and 2018Houston-demonstrate that MMCTNet achieves superior performance over existing methods in terms of overall accuracy (OA), average accuracy (AA), and the Kappa coefficient.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

MMCTNet: multimodal cross-scale transformer network for hyperspectral and LiDAR/SAR image classification — 科研速览 Science Skim