Scholay

学术搜索 · AI 审稿 · LaTeX 协作

The Actual Performance of ML/AI Models in Predicting Radiation-Induced Toxicity in Head and Neck Cancer: A Systematic Review and Meta-Analysis.

作者:Gibson C. Ugwu, Farzad Jalali, Geoffrey Liu, Guojun Li, Johannes A. Langendijk, Behrooz Z. Alizadeh · 发表于:Radiotherapy and Oncology · 年份:2025 · DOI:10.1016/j.radonc.2025.111350 · 被引用次数:2 · 研究领域:Radiomics and Machine Learning in Medical Imaging、Head and Neck Cancer Studies、Effects of Radiation Exposure

An increasing number of Artificial intelligence (AI) and machine learning (ML) models are being developed to predict radiation-induced toxicities (RITs) in patients with head and neck cancer (HNC). But their performance and reliability remain uncertain. This systematic review and meta-analysis evaluated the predictive accuracy and methodological quality of these models. We comprehensively searched PubMed, EMBASE, Web of Science, and the Cochrane Library to identify studies reporting on ML/AI models for predicting RITs in HNC patients. Eligible studies were assessed for bias risk using the PROBAST tool, and key performance metrics, including the area under the receiver operating curve (AUROC), were extracted. A hierarchical multilevel meta-analysis was performed to estimate pooled AUROC values, and subgroup analyses explored the influence of study characteristics on model performance. A total of 67 studies with a total of 568 models were included, showing moderate discriminatory power of ML/AI models, with a pooled AUROC = 0.76; 95 % CI: 0.73-0.78. Nonetheless, substantial heterogeneity was observed across studies. Incorporating imaging biomarkers significantly improved model performance. Prospective and internal validation showed comparable performance; external validation shows true generalizability. The predominance of retrospective designs and variability in predictor selection may have introduced bias, affecting model reliability and generalisability. ML/AI models hold pr...