"Dr. AI Will See You Now": How Do ChatGPT-4 Treatment Recommendations Align With Orthopaedic Clinical Practice Guidelines?

Tanios Dagher,Emma P Dwyer,Hayden P Baker,Senthooran Kalidoss,Jason A Strelzow

doi:10.1097/corr.0000000000003234

Abstract

Artificial intelligence (AI) is engineered to emulate tasks that have historically required human interaction and intellect, including learning, pattern recognition, decision-making, and problem-solving. Although AI models like ChatGPT-4 have demonstrated satisfactory performance on medical licensing exams, suggesting a potential for supporting medical diagnostics and decision-making, no study of which we are aware has evaluated the ability of these tools to make treatment recommendations when given clinical vignettes and representative medical imaging of common orthopaedic conditions. As AI continues to advance, a thorough understanding of its strengths and limitations is necessary to inform safe and helpful integration into medical practice. (1) What is the concordance between ChatGPT-4-generated treatment recommendations for common orthopaedic conditions with both the American Academy of Orthopaedic Surgeons (AAOS) clinical practice guidelines (CPGs) and an orthopaedic attending physician's treatment plan? (2) In what specific areas do the ChatGPT-4-generated treatment recommendations diverge from the AAOS CPGs? Ten common orthopaedic conditions with associated AAOS CPGs were identified: carpal tunnel syndrome, distal radius fracture, glenohumeral joint osteoarthritis, rotator cuff injury, clavicle fracture, hip fracture, hip osteoarthritis, knee osteoarthritis, ACL injury, and acute Achilles rupture. For each condition, the medical records of 10 deidentified patients managed at our facility were used to construct clinical vignettes that each had an isolated, single diagnosis with adequate clarity. The vignettes also encompassed a range of diagnostic severity to evaluate more thoroughly adherence to the treatment guidelines outlined by the AAOS. These clinical vignettes were presented alongside representative radiographic imaging. The model was prompted to provide a single treatment plan recommendation. Each treatment plan was compared with established AAOS CPGs and to the treatment plan documented by the attending orthopaedic surgeon treating the specific patient. Vignettes where ChatGPT-4 recommendations diverged from CPGs were reviewed to identify patterns of error and summarized. ChatGPT-4 provided treatment recommendations in accordance with the AAOS CPGs in 90% (90 of 100) of clinical vignettes. Concordance between ChatGPT-generated plans and the plan recommended by the treating orthopaedic attending physician was 78% (78 of 100). One hundred percent (30 of 30) of ChatGPT-4 recommendations for fracture vignettes and hip and knee arthritis vignettes matched with CPG recommendations, whereas the model struggled most with recommendations for carpal tunnel syndrome (3 of 10 instances demonstrated discordance). ChatGPT-4 recommendations diverged from AAOS CPGs for three carpal tunnel syndrome vignettes; two ACL injury, rotator cuff injury, and glenohumeral joint osteoarthritis vignettes; as well as one acute Achilles rupture vignette. In these situations, ChatGPT-4 most often struggled to correctly interpret injury severity and progression, incorporate patient factors (such as lifestyle or comorbidities) into decision-making, and recognize a contraindication to surgery. ChatGPT-4 can generate accurate treatment plans aligned with CPGs but can also make mistakes when it is required to integrate multiple patient factors into decision-making and understand disease severity and progression. Physicians must critically assess the full clinical picture when using AI tools to support their decision-making. ChatGPT-4 may be used as an on-demand diagnostic companion, but patient-centered decision-making should continue to remain in the hands of the physician.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

"Dr. AI Will See You Now": How Do ChatGPT-4 Treatment Recommendations Align With Orthopaedic Clinical Practice Guidelines?

Abstract

Talk to us

Similar Papers

More From: Clinical orthopaedics and related research

Lead the way for us

Similar Papers

Currently Available Large Language Models Do Not Provide Musculoskeletal Treatment Recommendations That Are Concordant With Evidence-Based Clinical Practice Guidelines
Benedict U Nwachukwu ... Kyle N Kunze
Arthroscopy: The Journal of Arthroscopic and Related Surgery | VOL. -
Benedict U Nwachukwu, et. al.Benedict U Nwachukwu ... Kyle N Kunze
01 Aug 2024
Arthroscopy: The Journal of Arthroscopic and Related Surgery | VOL. -

Linking of the American Academy of Orthopaedic Surgeons Distal Radius Fracture Clinical Practice Guidelines to the International Classification of Functioning, Disability, and Health; International Classification of Diseases; and ICF Core Sets for Hand Conditions
Saravanan Esakki ... Saipriya Vajravelu
HAND | VOL. 11
Saravanan Esakki, et. al.Saravanan Esakki ... Saipriya Vajravelu
07 Jul 2016
HAND | VOL. 11

Carpal Tunnel Syndrome Diagnosis and Treatment: A Survey of Members of the American Society for Surgery of the Hand
Lewis B Lane ... Nina Kohn
The Journal of Hand Surgery | VOL. 39
Lewis B Lane, et. al.Lewis B Lane ... Nina Kohn
13 Sep 2014
The Journal of Hand Surgery | VOL. 39

Agreement among ASES members on the AAOS Clinical Practice Guidelines.
E. Scott Paxton ... Joseph A. Abboud
Orthopedics | VOL. 38
E. Scott Paxton, et. al.E. Scott Paxton ... Joseph A. Abboud
01 Mar 2015
Orthopedics | VOL. 38

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

"Dr. AI Will See You Now": How Do ChatGPT-4 Treatment Recommendations Align With Orthopaedic Clinical Practice Guidelines?

Abstract

Talk to us

Similar Papers

More From: Clinical orthopaedics and related research