Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Bots in white coats: are large language models the future of patient education? A multicenter cross-sectional analysis

作者:Ughur Aghamaliyev, Javad Karimbayli, Athanasios Zamparas, Florian Bösch, Michael Thomas, Thomas Schmidt, Christian Krautz, Christoph Kahlert, Sebastian Schölch, Martin K. Angele, Hanno Nieß, Markus Guba, Jens Werner, Matthias Ilmer, Bernhard W. Renz · 发表于:International Journal of Surgery · 年份:2025 · DOI:10.1097/js9.0000000000002250 · 被引用次数:9 · 研究领域:Artificial Intelligence in Healthcare and Education、AI in Service Interactions、Clinical Reasoning and Diagnostic Skills

OBJECTIVES: Every year, around 300 million surgeries are conducted worldwide, with an estimated 4.2 million deaths occurring within 30 days after surgery. Adequate patient education is crucial, but often falls short due to the stress patients experience before surgery. Large language models (LLMs) can significantly enhance this process by delivering thorough information and addressing patient concerns that might otherwise go unnoticed. MATERIAL AND METHODS: This cross-sectional study evaluated Chat Generative Pretrained Transformer-4o's audio-based responses to frequently asked questions (FAQs) regarding six general surgical procedures. Three experienced surgeons and two senior residents formulated seven general and three procedure-specific FAQs for both preoperative and postoperative situations, covering six surgical scenarios (major: pancreatic head resection, rectal resection, total gastrectomy; minor: cholecystectomy, Lichtenstein procedure, hemithyroidectomy). In total, 120 audio responses were generated, transcribed, and assessed by 11 surgeons from 6 different German university hospitals. RESULTS: ChatGPT-4o demonstrated strong performance, achieving an average score of 4.12/5 for accuracy, 4.46/5 for relevance, and 0.22/5 for potential harm across 120 questions. Postoperative responses surpassed preoperative ones in both accuracy and relevance, while also exhibiting lower potential for harm. Additionally, responses related to minor surgeries were minimal, but signific...