Mode
Text Size
Log in / Sign up

Deep learning models for glioma segmentation achieve high whole tumor accuracy with DSC 0.860New AI tools show high accuracy for mapping brain tumors before surgery

AI-generated summary of the cited source, checked by automated accuracy review. How we work

Key Takeaway
Consider that deep learning glioma segmentation models show high whole tumor accuracy but variable performance across tumor regions.

This is a systematic review and meta-analysis of deep learning models for glioma segmentation on preoperative MRI. The authors synthesized evidence from 88 identified models, with 36 included in quantitative analyses. The key finding is high whole tumor segmentation accuracy, with a pooled Dice Similarity Coefficient of 0.860 (95% CI 0.840-0.881). Performance was significantly lower for enhancing, non-enhancing, and tumor core delineation. Models using 3D and multiparametric MRI inputs consistently outperformed those without, and training on BraTS datasets was associated with higher performance. A trend toward improved accuracy for high-grade gliomas was not statistically significant. Training dataset size was not associated with performance. Multivariate meta-regression found only publication year independently predicted improved accuracy (beta=0.023, p=0.017). The authors note that no single factor explained the variability observed across studies and that factors driving heterogeneity are poorly understood. Practice relevance is limited; future studies should prioritize multivariable analyses to better define determinants of model performance and support application into everyday practice.

Doctors need precise maps of brain tumors before surgery to plan safe procedures. A large review looked at 36 different computer models designed to outline gliomas, a type of brain tumor, using MRI scans. These AI tools showed high accuracy for mapping the whole tumor. The average score was 0.860 on a standard measure of precision. This means the computer outlines matched the human expert outlines very well in most cases.

However, the tools struggled with specific parts of the tumor. They were significantly less accurate at defining the enhancing and non-enhancing areas. Using advanced MRI inputs like 3D images or multiple types of scans helped the models perform better. Training the AI on a specific public dataset also led to higher scores. Yet, simply having more training data did not automatically make the tools better.

The review found that the year a study was published predicted better accuracy. This suggests technology is improving over time. However, researchers could not find one single reason why some models worked better than others. The reasons for these differences are still unclear. While the tools show promise, more research is needed to understand exactly what makes them succeed or fail before they become standard in every hospital.

What this means for you:
AI maps brain tumors well overall, but training data and scan types matter for accuracy.

Study Details

Study typeMeta analysis
EvidenceLevel 1
PublishedJun 2026
View Original Abstract ↓
BACKGROUND: Accurate segmentation is central for the diagnosis and treatment of gliomas. Although manual segmentation remains the clinical standard, it is time-consuming and subject to inter-operator variability. In recent years, deep learning (DL) models have been developed to automate this process, offering scalable alternatives. Performance across these models remains variable, and the factors driving this heterogeneity are poorly understood. METHODS: Following PRISMA guidelines, databases were searched for studies reporting the performance of models preoperatively segmenting gliomas. Data regarding model and patient characteristics were extracted, and subgroup analyses along with mixed-effects meta-regressions were performed to identify factors linked to segmentation accuracy, as measured by the Dice Similarity Coefficient (DSC). RESULTS: 88 models were identified, of which 36 were included in quantitative analyses. Whole tumor segmentation demonstrated a general high accuracy (DSC 0.860, 95% CI 0.840-0.881), with significantly lower performance found in enhancing, non-enhancing, and tumor core delineation. Subgroup analyses found models using 3D and multiparametric MRI inputs consistently outperformed those that did not. Models trained on BraTS datasets were associated with higher performance compared to original institutional data. Segmentation of high-grade gliomas showed a trend toward improved accuracy but was not statistically significant. Training dataset size was not associated with segmentation performance. In multivariate meta-regression, only publication year independently predicted improved accuracy (β=0.023, p = 0.017). CONCLUSION: Segmentation performance in DL-based glioma MRI is most consistently associated with the use of 3D model architectures and multiparametric MRI inputs. Models trained on BraTS datasets showed a trend toward higher performance, suggesting a possible benchmarking effect. However, in both univariate and multivariate analyses we found no single factor explained the variability observed across studies. Future studies should prioritize multivariable analyses to better define determinants of model performance and in turn support the application of these models into everyday practice.
Free Newsletter

Clinical research that matters. Delivered to your inbox.

Join thousands of clinicians and researchers. No spam, unsubscribe anytime.