Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Enhancing personalized anesthesia plans in cardiac surgery with AI: ChatGPT's advantages and the imperative for clinical oversight

作者:Zheng Chen, Yuxin Gao, Shiyao Gu, Zhengping Yong, Lei Li, Ruixuan Wang, Qian Lei, Si Zeng · 发表于:Anesthesiology and Perioperative Science · 年份:2026 · DOI:10.1007/s44254-025-00140-3 · 被引用次数:2 · 研究领域:Artificial Intelligence in Healthcare and Education、Machine Learning in Healthcare、Topic Modeling

Abstract Purpose To compare ChatGPT, a general-purpose large language model (LLM), with OpenBioLLM, a domain-specific biomedical model, in generating clinically appropriate anesthesia plans, and to assess the impact of advanced prompt engineering (PE). Methods A comparative observational study analyzing anonymized clinical records of 100 cardiac surgery patients using three LLMs under simple querying and PE. Plans were evaluated by anesthesiologists using a double-blind approach. The main outcome measures included clinical alignment, reasoning quality, medication selection, dosage accuracy, omissions, and risk of harm. Scores were rated on a 5-point Likert scale by both experienced and trainee anesthesiologists, and differences analyzed via paired t-tests and analysis of variance. Results In clinical matching accuracy, logical reasoning, medication selection, and safety evaluation metrics, ChatGPT consistently outperformed OpenBioLLM. GPT-4o demonstrated superior performance compared to GPT-3.5, with significant performance enhancements achieved through prompt optimization. Both physician groups reported notable improvements in their average scores for the ChatGPT model when utilizing advanced PE, whereas OpenBioLLM showed comparatively smaller score improvements. Conclusion ChatGPT’s adaptability and clinical accuracy, enhanced by PE, make it a valuable tool for anesthesia planning, especially in resource-limited settings. However, over-reliance on AI by less experienced cli...