Can AI Think Like a Plastic Surgeon? Evaluating GPT-4’s Clinical Judgment in Reconstructive Procedures of the Upper Extremity-Reference-Cited by-同舟云学术

Can AI Think Like a Plastic Surgeon? Evaluating GPT-4’s Clinical Judgment in Reconstructive Procedures of the Upper Extremity

Published:2023-12 Issue:12 Volume:11 Page:e5471
ISSN:2169-7574
Container-title:Plastic and Reconstructive Surgery - Global Open
language:en
Short-container-title:

Author:

Leypold Tim¹,Schäfer Benedikt¹,Boos Anja¹,Beier Justus P.¹

Affiliation:

1. From the Department of Plastic Surgery, Hand Surgery–Burn Center, University Hospital RWTH Aachen, Aachen, Germany.

Abstract

Summary: This study delves into the potential application of OpenAI’s Generative Pretrained Transformer 4 (GPT-4) in plastic surgery, with a particular focus on procedures involving the hand and arm. GPT-4, a cutting-edge artificial intelligence (AI) model known for its advanced chat interface, was tested on nine surgical scenarios of varying complexity. To optimize the performance of GPT-4, prompt engineering techniques were used to guide the model’s responses and improve the relevance and accuracy of its output. A panel of expert plastic surgeons evaluated the responses using a Likert scale to assess the model’s performance, based on five distinct criteria. Each criterion was scored on a scale of 1 to 5, with 5 representing the highest possible score. GPT-4 demonstrated a high level of performance, achieving an average score of 4.34 across all cases, consistent across different complexities. The study highlights the ability of GPT-4 to understand and respond to complicated surgical scenarios. However, the study also identifies potential areas for improvement. These include refining the prompts used to elicit responses from the model and providing targeted training with specialized, up-to-date sources. This study demonstrates a new approach to exploring large language models and highlights potential future applications of AI. These could improve patient care, refine surgical outcomes, and even change the way we approach complex clinical scenarios in plastic surgery. However, the intrinsic limitations of AI in its current state, together with the potential ethical considerations and the inherent uncertainty of unanticipated issues, serve to reiterate the indispensable role and unparalleled value of human plastic surgeons.

Publisher

Ovid Technologies (Wolters Kluwer Health)

Subject

Surgery,General Medicine

Reference9 articles.

1. Benefits, limits, and risks of GPT-4 as an AI Chatbot for Medicine.;Lee;N Engl J Med,2023

2. Application of ChatGPT in cosmetic plastic surgery: ally or antagonist?;Gupta;Aesthet Surg J,2023

3. Let’s chat about chatbots: additional thoughts on ChatGPT and its role in plastic surgery along with its ability to perform systematic reviews.;Najafali;Aesthet Surg J,2023

4. Expanding cosmetic plastic surgery research with ChatGPT.;Gupta;Aesthet Surg J,2023

5. Evaluation of the artificial intelligence chatbot on breast reconstruction and its efficacy in surgical research: a case study.;Xie;Aesthetic Plast Surg,2023

Cited by 9 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Accuracy assessment of ChatGPT responses to frequently asked questions regarding anterior cruciate ligament surgery;The Knee;2024-12

2. Prompt Engineering Paradigms for Medical Applications: Scoping Review;Journal of Medical Internet Research;2024-09-10

3. Artificial intelligence as an adjunctive tool in hand and wrist surgery: a review;Artificial Intelligence Surgery;2024-09-02

4. Large Language Models for Intraoperative Decision Support in Plastic Surgery: A Comparison between ChatGPT-4 and Gemini;Medicina;2024-06-08

5. Integrating AI in Lipedema Management: Assessing the Efficacy of GPT-4 as a Consultation Assistant;Life;2024-05-20