Artificial Intelligence Models Are Limited in Predicting Clinical Outcomes Following Hip Arthroscopy-Reference-Cited by-同舟云学术

Artificial Intelligence Models Are Limited in Predicting Clinical Outcomes Following Hip Arthroscopy

Published:2024-08 Issue:8 Volume:12 Page:
ISSN:2329-9185
Container-title:JBJS Reviews
language:en
Short-container-title:

Author:

Mehta Apoorva¹^ORCID,El-Najjar Dany¹^ORCID,Howell Harrison¹^ORCID,Gupta Puneet¹^ORCID,Arciero Emily¹^ORCID,Marigi Erick M.²^ORCID,Parisien Robert L.³^ORCID,Trofa David P.¹^ORCID

Affiliation:

1. Department of Orthopaedic Surgery, Columbia University Irving Medical Center, New York, New York

2. Department of Orthopaedic Surgery, Mayo Clinic, Rochester, Minnesota

3. Department of Orthopaedic Surgery, Mount Sinai, New York, New York

Abstract

Background: Hip arthroscopy has seen a significant surge in utilization, but complications remain, and optimal functional outcomes are not guaranteed. Artificial intelligence (AI) has emerged as an effective supportive decision-making tool for surgeons. The purpose of this systematic review was to characterize the outcomes, performance, and validity (generalizability) of AI-based prediction models for hip arthroscopy in current literature. Methods: Two reviewers independently completed structured searches using PubMed/MEDLINE and Embase databases on August 10, 2022. The search query used the terms as follows: (artificial intelligence OR machine learning OR deep learning) AND (hip arthroscopy). Studies that investigated AI-based risk prediction models in hip arthroscopy were included. The primary outcomes of interest were the variable(s) predicted by the models, best model performance achieved (primarily based on area under the curve, but also accuracy, etc), and whether the model(s) had been externally validated (generalizable). Results: Seventy-seven studies were identified from the primary search. Thirteen studies were included in the final analysis. Six studies (n = 6,568) applied AI for predicting the achievement of minimal clinically important difference for various patient-reported outcome measures such as the visual analog scale and the International Hip Outcome Tool 12-Item Questionnaire, with area under a receiver-operating characteristic curve (AUC) values ranging from 0.572 to 0.94. Three studies used AI for predicting repeat hip surgery with AUC values between 0.67 and 0.848. Four studies focused on predicting other risks, such as prolonged postoperative opioid use, with AUC values ranging from 0.71 to 0.76. None of the 13 studies assessed the generalizability of their models through external validation. Conclusion: AI is being investigated for predicting clinical outcomes after hip arthroscopy. However, the performance of AI models varies widely, with AUC values ranging from 0.572 to 0.94. Critically, none of the models have undergone external validation, limiting their clinical applicability. Further research is needed to improve model performance and ensure generalizability before these tools can be reliably integrated into patient care. Level of Evidence: Level IV. See Instructions for Authors for a complete description of levels of evidence.

Publisher

Ovid Technologies (Wolters Kluwer Health)

Reference46 articles.

1. A shift in hip arthroscopy use by patient age and surgeon volume: A New York state-based population analysis 2004 to 2016;Schairer;Arthroscopy,2019

2. Indications for hip arthroscopy;Ross;Sports Health,2017

3. Does hip preservation surgery prevent arthroplasty? Quantifying the rate of conversion to arthroplasty following hip preservation surgery;Sohatee;J Hip Preserv Surg,2020

4. Predictors of poor clinical outcome after arthroscopic labral preservation, capsular plication, and cam osteoplasty in the setting of borderline hip dysplasia;Hatakeyama;Am J Sports Med,2018

5. Preoperative symptom duration is associated with outcomes after hip arthroscopy;Basques;Am J Sports Med,2019