Bibliographic Details
| Title: |
Gender bias in resident evaluations: Natural language processing and competency evaluation. |
| Authors: |
Andrews, Jane, Chartash, David, Hay, Seonaid |
| Source: |
Medical Education. Dec2021, Vol. 55 Issue 12, p1383-1342. 6p. |
| Subjects: |
National competency-based educational tests, Teacher-student relationships, Hospital medical staff, Internal medicine, Natural language processing, Sex discrimination, Thematic analysis, Job performance |
| Abstract: |
Background: Research shows that female trainees experience evaluation penalties for gender non- conforming behaviour during medical training. Studies of medical education evaluations and performance scores do reflect a gender bias, though studies are of varying methodology and results have not been consistent. Objective: We sought to examine the differences in word use, competency themes and length within written evaluations of internal medicine residents at scale, considering the impact of both faculty and resident gender. We hypothesised that female internal medicine residents receive more negative feedback, and different thematic feedback than male residents. Methods: This study utilised a corpus of 3864 individual responses to positive and negative questions over the course of six years (2012- 2018) within Yale University School of Medicine's internal medicine residency. Researchers developed a sentiment model to assess the valence of evaluation responses. We then used natural language processing (NLP) to evaluate whether female versus male residents received more positive or negative feedback and if that feedback focussed on different Accreditation Council for Graduate Medical Education (ACGME) core competencies based on their gender. Evaluator- evaluatee gender dyad was analysed to see how it impacted quantity and quality of feedback. Results: We found that female and male residents did not have substantively different numbers of positive or negative comments. While certain competencies were discussed more than others, gender did not seem to influence which competencies were discussed. Neither gender trainee received more written feedback, though female evaluators tended to write longer evaluations. Conclusions: We conclude that when examined at scale, quantitative gender differences are not as prevalent as has been seen in qualitative work. We suggest that further investigation of linguistic phenomena (such as context) is warranted to reconcile this finding with prior work. [ABSTRACT FROM AUTHOR] |
|
Copyright of Medical Education is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.) |
| Database: |
Psychology and Behavioral Sciences Collection |