Scholay

学术搜索 · AI 审稿 · LaTeX 协作

The Large Language Model ChatGPT-4 Demonstrates Excellent Triage Capabilities and Diagnostic Performance for Patients Presenting with Various Causes of Knee Pain.

作者:K. Kunze, Nathan H. Varady, Michael Mazzucco, Amy Z. Lu, J. Chahla, R. Martin, Anil S. Ranawat, Andrew D. Pearle, Riley J. Williams · 发表于:Arthroscopy: The Journal of Arthroscopy And Related · 年份:2024 · DOI:10.1016/j.arthro.2024.06.021 · 被引用次数:38 · 研究领域:Medicine

PURPOSE To provide a proof-of-concept analysis of the appropriateness and performance of ChatGPT-4 to triage, synthesize differential diagnoses, and generate treatment plans concerning common presentations of knee pain. METHODS Twenty knee complaints warranting triage and expanded scenarios were input into ChatGPT-4, with memory cleared prior to each new input to mitigate bias. For the 10 triage complaints, ChatGPT-4 was asked to generate a differential diagnosis which was graded for accuracy and suitability in comparison to a differential created by two orthopaedic sports medicine physicians. For the 10 clinical scenarios, ChatGPT-4 was prompted to provide treatment guidance for the patient, which was again graded. To test the higher-order capabilities of ChatGPT-4, further inquiry into these specific management recommendations was performed and graded. RESULTS All ChatGPT-4 diagnoses were deemed appropriate within the spectrum of potential pathologies on a differential. The top diagnosis on the differential was identical between surgeons and ChatGPT-4 for 70% of scenarios, and the top diagnosis provided by the surgeon appeared as either the first or second diagnosis in 90% of scenarios. Overall, 16/30 (53.3%) of diagnoses in the differential were identical. When provided with 10 expanded vignettes with a single diagnosis, the accuracy of ChatGPT-4 increased to 100%, with the suitability of management graded as appropriate in 90% of cases. Specific information pertaining...