Skip to content
Open access

LMCAN: A Lightweight Multiscale Contextual Attention Network for Hyperspectral and Multispectral Image Fusion

Sep 2026 · Remote Sensing · 0 citations · 44 references

Abstract

Hyperspectral and multispectral image fusion requires enhancing spatial details while preserving the pixel-wise spectral fidelity of hyperspectral observations. Transformer-based approaches provide effective contextual modeling, yet hierarchical token merging may cause spectral mixing, whereas pixel-preserving tokenization restricts the spatial context captured by window attention. To resolve this trade-off, we propose a Lightweight Multiscale Contextual Attention Network (LMCAN) that reconstructs spatial context without altering the pixel-level representation. The proposed network progressively broadens contextual perception and coordinates complementary spatial dependencies within the attention process, enabling information at different scales to interact adaptively rather than being modeled in isolation. It further recovers local correlations omitted by fixed window partitioning, improving the reconstruction of boundaries and fine structures. Through this unified design, spectral preservation and spatial context restoration are jointly achieved within a shallow and efficient architecture. Experiments on four benchmark datasets demonstrate competitive accuracy with substantially lower complexity. On CAVE, LMCAN achieves 49.04 dB PSNR and 2.47 SAM with only 0.157 M parameters and 12.06 G FLOPs.

Read PDF

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.