Enhancing MR image super-resolution with a multi-head attention network: capturing high-dimensional medical features
摘要
Magnetic resonance (MR) imaging is a cornerstone of modern clinical diagnosis, yet its quality is often limited by acquisition constraints, potentially obscuring subtle pathological details and hindering diagnostic accuracy. To address this, we propose a novel multi-head attention network (MHAN) for MR image super-resolution, specifically engineered to capture and reconstruct high-dimensional medical features. The core of MHAN features two key innovations: a 3D multi-head attention module (3D-MHead) and a depth–width feature fusion module (DW-FFM). The 3D-MHead employs multi-scale 3D convolutions to dynamically recalibrate feature weights, enhancing the model’s focus on critical high-frequency details such as lesion boundaries. Concurrently, the DW-FFM effectively integrates shallow spatial textures with deep semantic information, ensuring both global structural coherence and local feature fidelity in the final output. Our extensive experiments on public MR datasets, such as IXI, demonstrate that MHAN significantly outperforms state-of-the-art methods in both quantitative metrics (PSNR/SSIM) and qualitative visual assessment. For instance, on the IXI-PD dataset at a