Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Digitizing Diagnoses: Distinguishing Infantile Hemangiomas From Other Vascular Anomalies

作者:Aretha On, Elena Huang, Jessica Hills, Jared Pasternak, Joanne Alfandre, Griffin Stockton Hogrogian, Albert C. Yan · 发表于:Pediatric Dermatology · 年份:2025 · DOI:10.1111/pde.70008 · 被引用次数:3 · 研究领域:Vascular Malformations and Hemangiomas、Central Venous Catheters and Hemodialysis、Vascular Malformations Diagnosis and Treatment

BACKGROUND/OBJECTIVES: Generative artificial intelligence (AI) models have become increasingly accessible and advanced with multimodal input. Infantile hemangiomas (IHs) are the most common pediatric vascular tumor, but pediatricians may have difficulty distinguishing them from similar-appearing lesions. We assess the capability of a public AI model, ChatGPT 4.0 (GPT) to distinguish between IHs and other vascular anomalies (VAs), evaluating its potential role as a clinical tool for pediatric clinicians. METHODS: This retrospective study assessed 50 IH and 50 non-IH VA images using a GPT zero-shot approach with the binary task of diagnosing an IH (or not). The same images were provided to four general pediatricians; comparison was performed between pediatricians and GPT. RESULTS: GPT achieves 75% accuracy in IH identification with an F1 score of 0.742. ROC curve generation yields an AUC of 0.80. Hundred-image analysis demonstrates that actual diagnosis affects GPT accuracy (p = 0.015). 50-IH-image analysis reveals configuration (p = 0.027), skin phototype (p = 0.019), and anatomical location (p = 0.047) as factors that may affect GPT accuracy. Comparing GPT to pediatricians reveals comparable results (p = 0.345). CONCLUSION: This off-the-shelf publicly available GPT (75%) was less accurate than a previously published well-trained AI that achieved higher accuracy (~92%). GPT's F1 score of 0.742 indicates moderate balance between precision and sensitivity, and an AUC calculation...