An increasing number of convolutional neural networks for fracture recognition and classification in orthopaedics-Reference-Cited by-同舟云学术

An increasing number of convolutional neural networks for fracture recognition and classification in orthopaedics

Published:2021-10-01 Issue:10 Volume:2 Page:879-885
ISSN:2633-1462
Container-title:Bone & Joint Open
language:en
Short-container-title:Bone & Joint Open

Author:

Oliveira e Carmo Luisa¹,van den Merkhof Anke²³,Olczak Jakub⁴,Gordon Max⁴,Jutte Paul C.¹,Jaarsma Ruurd L.²³,IJpma Frank F. A.⁵,Doornberg Job N.¹²³⁵,Prijs Jasper¹²³⁵^ORCID,

Affiliation:

1. Department of Orthopaedic Surgery, University Medical Centre, University of Groningen, Groningen, Groningen, Netherlands

2. Department of Orthopaedic Surgery, Flinders Medical Centre, Bedford Park, Adelaide, South Australia, Australia

3. Flinders University, Bedford Park, Adelaide, South Australia, Australia

4. Institute of Clinical Sciences, Danderyd University Hospital, Karolinska Institute, Stockholm, Sweden

5. Department of Trauma Surgery, University Medical Centre Groningen, University of Groningen, Groningen, Groningen, Netherlands

Abstract

Aims The number of convolutional neural networks (CNN) available for fracture detection and classification is rapidly increasing. External validation of a CNN on a temporally separate (separated by time) or geographically separate (separated by location) dataset is crucial to assess generalizability of the CNN before application to clinical practice in other institutions. We aimed to answer the following questions: are current CNNs for fracture recognition externally valid?; which methods are applied for external validation (EV)?; and, what are reported performances of the EV sets compared to the internal validation (IV) sets of these CNNs? Methods The PubMed and Embase databases were systematically searched from January 2010 to October 2020 according to the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) statement. The type of EV, characteristics of the external dataset, and diagnostic performance characteristics on the IV and EV datasets were collected and compared. Quality assessment was conducted using a seven-item checklist based on a modified Methodologic Index for NOn-Randomized Studies instrument (MINORS). Results Out of 1,349 studies, 36 reported development of a CNN for fracture detection and/or classification. Of these, only four (11%) reported a form of EV. One study used temporal EV, one conducted both temporal and geographical EV, and two used geographical EV. When comparing the CNN’s performance on the IV set versus the EV set, the following were found: AUCs of 0.967 (IV) versus 0.975 (EV), 0.976 (IV) versus 0.985 to 0.992 (EV), 0.93 to 0.96 (IV) versus 0.80 to 0.89 (EV), and F1-scores of 0.856 to 0.863 (IV) versus 0.757 to 0.840 (EV). Conclusion The number of externally validated CNNs in orthopaedic trauma for fracture recognition is still scarce. This greatly limits the potential for transfer of these CNNs from the developing institute to another hospital to achieve similar diagnostic performance. We recommend the use of geographical EV and statements such as the Consolidated Standards of Reporting Trials–Artificial Intelligence (CONSORT-AI), the Standard Protocol Items: Recommendations for Interventional Trials–Artificial Intelligence (SPIRIT-AI) and the Transparent Reporting of a multivariable prediction model for Individual Prognosis or Diagnosis–Machine Learning (TRIPOD-ML) to critically appraise performance of CNNs and improve methodological rigor, quality of future models, and facilitate eventual implementation in clinical practice. Cite this article: Bone Jt Open 2021;2(10):879–885.

Publisher

British Editorial Society of Bone & Joint Surgery

Subject

Pharmacology (medical),Complementary and alternative medicine,Pharmaceutical Science

Link

https://online.boneandjoint.org.uk/doi/pdf/10.1302/2633-1462.210.BJO-2021-0133

Reference48 articles.

1. High-performance medicine: the convergence of human and artificial intelligence

2. Current Applications and Future Impact of Machine Learning in Radiology

3. Reporting guidelines for clinical trials evaluating artificial intelligence interventions are needed

4. Reporting of artificial intelligence prediction models

5. Computer vs human: Deep learning versus perceptual training for the detection of neck of femur fractures

Cited by 28 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. AFFnet - a deep convolutional neural network for the detection of atypical femur fractures from anteriorposterior radiographs;Bone;2024-10

2. AI for detection, classification and prediction of loss of alignment of distal radius fractures; a systematic review;European Journal of Trauma and Emergency Surgery;2024-07-09

3. Artificial intelligence in musculoskeletal imaging: realistic clinical applications in the next decade;Skeletal Radiology;2024-06-20

4. Predicting factors for extremity fracture among border-fall patients using machine learning computing;Heliyon;2024-06

5. A Review on the Use of Artificial Intelligence in Fracture Detection;Cureus;2024-04-16