Skip to content

A Deep Learning-Driven Brain Tumour Segmentation using a Hybrid U-Net and LSTM Architecture

Jul 2026 · Pertanika journal of science & technology · 0 citations

TL;DR

A hybrid deep learning framework consisting of a U-Net architecture integrated with Long Short-Term Memory (LSTM) networks is designed, which displays the promise of combining a convolutional and a recurrent architecture to propagate automated neuroimaging analysis.

Abstract

The most important step in diagnosis, treatment planning, and prediction analysis is the accurate segmentation of brain tumours from magnetic resonance imaging (MRI) scans. Radiologists' manual segmentation was time-consuming, subjective, and inaccurate. For this reason, there is a need for automated approaches. In Current days, deep learning (DL) has shown great promise for enhancing the processing of medical images. In DL, the U-Net architecture has become a typical framework for medical image segmentation of images due to its symmetric encoder–decoder design and skip links that preserve spatial detail. At the same time, conventional U-Net models are restricted to 2D slices and cannot detect contextual connections between slices in spatial MRI images; this may lead to discontinuities and reduced accuracy. The study suggests addressing the above-mentioned challenges, a hybrid deep learning framework consisting of a U-Net architecture integrated with Long Short-Term Memory (LSTM) networks is designed. The U-Net component extracts the variant-invariant spatial and structural features, and LSTM is responsible for integrating the temporal dependencies spatially, which inject discontinuities across parallel slices, which is imperative for the boundary delineation and localisation of the tumour. Compared to the baseline model, the suggested hybrid U-Net and LSTM networks exhibit noticeably better segmentation accuracy and visual consistency under various scenarios after being trained and assessed on a publicly accessible brain MRI dataset. According to experimental data, the suggested model provides great segmentation performance, with Dice scores above 95% and accuracy above 96%. To sum up, the whole exercise displays the promise of combining a convolutional and a recurrent architecture to propagate automated neuroimaging analysis. This will not only reduce the manual labour but also continue to work well in clinical practice, as it is comfortable and prepared for replication in the future.

View source

Similar papers

Open access Sep 2026

BRAIN TUMOR SEGMENTATION OF MRI SEQUENCES (T1, T2, T1CE, FLAIR) USING BRATS DATASET

Brain tumor segmentation from Magnetic Resonance Imaging (MRI) is an important task in computer-aided diagnosis because accurate identification of tumor regions supports clinical assessment and treatment planning. However, the complex structure, irregular shape, intensity variation, and heterogeneous appearance of brain tumors make automated segmentation challenging. This study presents a comparative deep learning framework for brain tumor segmentation using the BraTS 2020 dataset and two-dimensional (2D) MRI images. In this study, we explore 2D deep learning architectures for automated tumor segmentation, focusing on U-Net, Vision Transformer (ViT), and Res-ViT models. U-Net, with its encoder–decoder design and skip connections, has been widely adopted for medical image segmentation due to its ability to capture fine-grained spatial features. ViT, leveraging self-attention mechanisms, introduces a global receptive field that enhances contextual understanding across slices. The Res-ViT hybrid combines residual learning with transformer-based attention, aiming to balance local feature extraction and long-range dependency modelling. Preprocessing steps, including skull stripping, intensity normalisation, and bias field correction, were applied to ensure consistency across scans. Data augmentation techniques such as rotation, flipping, and elastic deformation were employed to mitigate overfitting and improve generalisation. The models are evaluated using important segmentation metrics, including Intersection over Union (IoU), accuracy, precision, recall/sensitivity, loss, and 95th-percentile Hausdorff Distance (HD95). The comparative analysis aims to identify the strengths and limitations of convolutional and transformer-based approaches for 2D brain tumor segmentation. The study demonstrates the potential of combining local feature extraction and global contextual learning to achieve more accurate and robust brain tumor segmentation from multimodal MRI images. Keywords: MRI; U-Net; ViT; BraTS 2020; IOU; HD95.

Lovedeep Kaur, Parminder Singh, Naveen Dhillon · 0 citations
Open access Jul 2026

Multimodal Brain Tumour Classification and Segmentation Using a Dual-Attention Swin-UNet with Quantum-Inspired Optimisation for MRI-Based Diagnostics

A Multimodal Brain Tumour Classification and Segmentation framework based on a Dual-Attention Swin-UNet architecture, enhanced with a Quantum-Inspired Optimisation (QIO) technique, to overcome limitations in transformer-based medical imaging models.

Anirban Mondal, V. E. Jesi · 0 citations
Open access 2026

Accurate Brain Tumor Classification Using MRI Images Based on A Hybrid Vision Transformer and BiLSTM Framework

Results show that ViT–BiLSTM's classification performance is superior to those of traditional deep learning methods: among all the tumor categories its accuracy is higher, its fine-tuning more perfect, as well as, its Recall rates greater.

Nagham Salim Mohammed, Omar S. Almolaa, A. S. Abdullah et al. · 0 citations
Conference Jul 2026

A Dual-Attention and Uncertainty-Aware Deep Learning Framework for Clinically Trusted Brain Tumor Segmentation

Although brain tumour segmentation is essential for medical image analysis, reliable and accurate segmentation with clinical and routine quality appears to be difficult because of the heterogeneity of the tumour, the lack of clear boundaries and the noise within the images. Current deep learning (DL) models have been plagued by the problem of the black box effect and overconfident predictions, limiting their use in real clinical settings. In fact, current DL models have experienced the issue of black box effect and high prediction accuracy, which prevent the models from being used in the real clinical setting. To solve these, the study puts forward a Dual-Attention and Uncertainty-Aware DL framework to enhance segmentation accuracy and clinical trustworthiness. The method combines both spatial and channel attention mechanism and is implemented in an encoder-decoder framework which allows obtaining more informative features from multi-modal MRI scans. Besides, a multi-scale feature fusion module is used to capture tumors of different scales, and an uncertainty-aware module is used to measure the confidence of prediction by computing the stochastic inference. A clinical trust calibration layer is added to make the predicted probabilities more realistic. Experiments performed on the BraTS dataset show better performance when compared with the state-of-the-art models with Dice score, IoU, sensitivity, and specificity of 94.8%, 89.5%, 93.6%, and 97.1%, respectively. The proposed framework not only enhances the segmentation accuracy but also generates the uncertainty maps, which are well suited for a clinical decision support system. Finally, the model provides a reliable, interpretable, and strong brain tumor segmentation solution, which has great potential in medical applications.

P. Prasana, V. Muneeswaran · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.