arrow
返回

Competency-Based Assessments: Leveraging Artificial Intelligence to Predict Subcompetency Content

delete2022-12-05
delete5
delete
OA
AI
G
Gregory J Booth *
B
Benjamin P. Ross
W
William A. Cronin
A
Angela McElrath
K
Kyle L. Cyr
J
John A. Hodgson
C
Charles G. Sibley
J
Johanes M. Ismawan
A
Alyssa Zuehl
J
James Slotto
M
Maureen Higgs
M
Matthew Haldeman
P
Phillip Geiger
D
Dink Jardine
DOI:10.1097/ACM.0000000000005115delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
PurposeFaculty feedback on trainees is critical to guiding trainee progress in a competency-based medical education framework. The authors aimed to develop and evaluate a Natural Language Processing (NLP) algorithm that automatically categorizes narrative feedback into corresponding Accreditation Council for Graduate Medical Education Milestone 2.0 subcompetencies.MethodTen academic anesthesiologists analyzed 5,935 narrative evaluations on anesthesiology trainees at 4 graduate medical education (GME) programs between July 1, 2019, and June 30, 2021. Each sentence (n = 25,714) was labeled with the Milestone 2.0 subcompetency that best captured its content or was labeled as demographic or not useful. Inter-rater agreement was assessed by Fleiss' Kappa. The authors trained an NLP model to predict feedback subcompetencies using data from 3 sites and evaluated its performance at a fourth site. Performance metrics included area under the receiver operating characteristic curve (AUC), positive predictive value, sensitivity, F1, and calibration curves. The model was implemented at 1 site in a self-assessment exercise.ResultsFleiss' Kappa for subcompetency agreement was moderate (0.44). Model performance was good for professionalism, interpersonal and communication skills, and practice-based learning and improvement (AUC 0.79, 0.79, and 0.75, respectively). Subcompetencies within medical knowledge and patient care ranged from fair to excellent (AUC 0.66-0.84 and 0.63-0.88, respectively). Performance for systems-based practice was poor (AUC 0.59). Performances for demographic and not useful categories were excellent (AUC 0.87 for both). In approximately 1 minute, the model interpreted several hundred evaluations and produced individual trainee reports with organized feedback to guide a self-assessment exercise. The model was built into a web-based application.ConclusionsThe authors developed an NLP model that recognized the feedback language of anesthesiologists across multiple GME programs. The model was operationalized in a self-assessment exercise. It is a powerful tool which rapidly organizes large amounts of narrative feedback.
Keyword:
FEEDBACK

期刊

Academic Medicine 封面图
Academic Medicine
IF:
5.2
论文数:
1.2W
被引数:
2.4W

机构

United States Navy 封面图
United States Navy
学者数:
6.7K
论文数: 5.5K
被引数: 175
United States Department of Defense 封面图
United States Department of Defense
学者数:
2.8W
论文数: 2.3W
被引数: 172
引用论文

引用论文

Assessment of Gender-Based Linguistic Differences in Physician Trainee Evaluations of Medical Faculty Using Automated Text Mining
err2019-05-10
err37
errOAAI
errHeath, Janae K.; Weissman, Gary E.; Clancy, Caitlin B.; Shou, Haochang; Farrar, John T.; Dine, C. Jessica
err分享
err收藏
err分享
err收藏
err分享
err收藏
Guidelines for Developing and Reporting Machine Learning Predictive Models in Biomedical Research: A Multidisciplinary View
err2016-12-16
err674
errOAAI
errLuo, Wei; Phung, Dinh; Tran, Truyen; Gupta, Sunil; Rana, Santu; Karmakar, Chandan; Shilton, Alistair; Yearwood, John; Dimitrova, Nevenka; Ho, Tu Bao; Venkatesh, Svetha; Berk, Michael
err分享
err收藏
Defects and Photonic Wells in One-Dimensional Photonic Lattices
err1996-12-15
err0
PREAI
errHiroshi Miyazaki; Yoji Jimba; Chong-Yeal Kim; Takeshi Watanabe
err分享
err收藏
err分享
err收藏
Using Natural Language Processing to Automatically Assess Feedback Quality: Findings From 3 Surgical Residencies
err2021-05-04
err21
errOAAI
errOtles, Erkin; Kendrick, Daniel E.; Solano, Quintin P.; Schuller, Mary; Ahle, Samantha L.; Eskender, Mickyas H.; Carnes, Emily; George, Brian C.
err分享
err收藏
学者 查看更多内容