Harnessing Natural Language Processing to Support Decisions Around Workplace-Based Assessment: Machine Learning Study of Competency-Based Medical Education-Reference-Cited by-同舟云学术

Harnessing Natural Language Processing to Support Decisions Around Workplace-Based Assessment: Machine Learning Study of Competency-Based Medical Education

Published:2022-05-27 Issue:2 Volume:8 Page:e30537
ISSN:2369-3762
Container-title:JMIR Medical Education
language:en
Short-container-title:JMIR Med Educ

Author:

Yilmaz Yusuf^ORCID,Jurado Nunez Alma^ORCID,Ariaeinejad Ali^ORCID,Lee Mark^ORCID,Sherbino Jonathan^ORCID,Chan Teresa M^ORCID

Abstract

Background Residents receive a numeric performance rating (eg, 1-7 scoring scale) along with a narrative (ie, qualitative) feedback based on their performance in each workplace-based assessment (WBA). Aggregated qualitative data from WBA can be overwhelming to process and fairly adjudicate as part of a global decision about learner competence. Current approaches with qualitative data require a human rater to maintain attention and appropriately weigh various data inputs within the constraints of working memory before rendering a global judgment of performance. Objective This study explores natural language processing (NLP) and machine learning (ML) applications for identifying trainees at risk using a large WBA narrative comment data set associated with numerical ratings. Methods NLP was performed retrospectively on a complete data set of narrative comments (ie, text-based feedback to residents based on their performance on a task) derived from WBAs completed by faculty members from multiple hospitals associated with a single, large, residency program at McMaster University, Canada. Narrative comments were vectorized to quantitative ratings using the bag-of-n-grams technique with 3 input types: unigram, bigrams, and trigrams. Supervised ML models using linear regression were trained with the quantitative ratings, performed binary classification, and output a prediction of whether a resident fell into the category of at risk or not at risk. Sensitivity, specificity, and accuracy metrics are reported. Results The database comprised 7199 unique direct observation assessments, containing both narrative comments and a rating between 3 and 7 in imbalanced distribution (scores 3-5: 726 ratings; and scores 6-7: 4871 ratings). A total of 141 unique raters from 5 different hospitals and 45 unique residents participated over the course of 5 academic years. When comparing the 3 different input types for diagnosing if a trainee would be rated low (ie, 1-5) or high (ie, 6 or 7), our accuracy for trigrams was 87%, bigrams 86%, and unigrams 82%. We also found that all 3 input types had better prediction accuracy when using a bimodal cut (eg, lower or higher) compared with predicting performance along the full 7-point rating scale (50%-52%). Conclusions The ML models can accurately identify underperforming residents via narrative comments provided for WBAs. The words generated in WBAs can be a worthy data set to augment human decisions for educators tasked with processing large volumes of narrative assessments.

Publisher

JMIR Publications Inc.

Subject

Computer Science Applications,Education

Reference56 articles.

1. Attending Emergency Physicians’ Perceptions of a Programmatic Workplace-Based Assessment System: The McMaster Modular Assessment Program (McMAP)

2. Workplace-based assessments in Wessex: the first 6 months

3. The McMaster Modular Assessment Program (McMAP)

4. ‘Playing the game’: How do surgical trainees seek feedback using workplace-based assessment?

5. The role of assessment in competency-based medical education

Cited by 14 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Leveraging Narrative Feedback in Programmatic Assessment: The Potential of Automated Text Analysis to Support Coaching and Decision-Making in Programmatic Assessment;Advances in Medical Education and Practice;2024-07

2. Validating the Physician Documentation Quality Instrument for Intensive Care Unit–Ward Transfer Notes;ATS Scholar;2024-06

3. The emerging role of generative artificial intelligence in transplant medicine;American Journal of Transplantation;2024-06

4. Demystifying AI: Current State and Future Role in Medical Education Assessment;ACAD MED;2024

5. Finding Medicine's Moneyball: How Lessons From Major League Baseball Can Advance Assessment in Precision Education;ACAD MED;2024