Effects of item and rater characteristics on checklist recording: what should we look for?

Saved in:
Bibliographic Details
Title: Effects of item and rater characteristics on checklist recording: what should we look for?
Authors: Huber, Philippe (AUTHOR), Baroffio, Anne (AUTHOR), Chamot, Eric (AUTHOR), Herrmann, François (AUTHOR), Nendaz, Mathieu R (AUTHOR), Vu, Nu V (AUTHOR)
Source: Medical Education. Aug2005, Vol. 39 Issue 8, p852-858. 7p.
Subjects: Medical students, Performance, Evaluation, Patients, Medical schools, Medical education
Abstract: Examinations based on using standardised patients (SPs) commonly use checklist recordings to evaluate students' clinical performance. This paper examines whether and to what extent item and rater characteristics affect the reliability of history checklist recording in an SP-based assessment. Checklist items were reviewed for the presence or absence of 5 item characteristics and a 2-point versus 3-point scoring scale. Agreement between checklist recordings obtained from SPs and clinician-examiners (CEs) were compared by item characteristics, scoring scale and CEs' level of involvement in the assessment. Based on 3179 pairs of recordings, the overall percentage of agreement between SPs and CEs was 83% (kappa = 0.64). Agreement was significantly higher for items scored on a 2-point than on a 3-point scale, and when the CE was also the author and the trainer of the station. After controlling for other factors, item characteristics were only marginally associated with level of interrater agreement. This study suggests that attention should be paid to specific aspects of checklist development and checklist recording training when an SP or CE is used as recorder. [ABSTRACT FROM AUTHOR]
Copyright of Medical Education is the property of Wiley-Blackwell and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Psychology and Behavioral Sciences Collection
Full text is not displayed to guests.
Description
Abstract:Examinations based on using standardised patients (SPs) commonly use checklist recordings to evaluate students' clinical performance. This paper examines whether and to what extent item and rater characteristics affect the reliability of history checklist recording in an SP-based assessment. Checklist items were reviewed for the presence or absence of 5 item characteristics and a 2-point versus 3-point scoring scale. Agreement between checklist recordings obtained from SPs and clinician-examiners (CEs) were compared by item characteristics, scoring scale and CEs' level of involvement in the assessment. Based on 3179 pairs of recordings, the overall percentage of agreement between SPs and CEs was 83% (kappa = 0.64). Agreement was significantly higher for items scored on a 2-point than on a 3-point scale, and when the CE was also the author and the trainer of the station. After controlling for other factors, item characteristics were only marginally associated with level of interrater agreement. This study suggests that attention should be paid to specific aspects of checklist development and checklist recording training when an SP or CE is used as recorder. [ABSTRACT FROM AUTHOR]
ISSN:03080110
DOI:10.1111/j.1365-2929.2005.02226.x