Effect of Rater Training on the Reliability of Technical Skill Assessments: a Randomized Controlled Trial
Overview
Affiliations
Background: Rater training improves the reliability of observational assessment tools but has not been well studied for technical skills. This study assessed whether rater training could improve the reliability of technical skill assessment.
Methods: Academic and community surgeons in Royal College of Physicians and Surgeons of Canada surgical subspecialties were randomly allocated to either rater training (7-minute video incorporating frame-of-reference training elements) or no training. Participants then assessed trainees performing a suturing and knot-tying task using 3 assessment tools: a visual analogue scale, a task-specific checklist and a modified version of the Objective Structured Assessment of Technical Skill global rating scale (GRS). We measured interrater reliability (IRR) using intraclass correlation type 2.
Results: There were 24 surgeons in the training group and 23 in the no-training group. Mean assessment tool scores were not significantly different between the 2 groups. The training group had higher IRR than the no-training group on the visual analogue scale (0.71 v. 0.46), task-specific checklist (0.46 v. 0.33) and GRS (0.71 v. 0.61). However, confidence intervals were wide and overlapping for all 3 tools.
Conclusion: For education purposes, the reliability of the visual analogue scale and GRS would be considered "good" for the training group but "moderate" for the no-training group. However, a significant difference in IRR was not shown, and reliability remained below the desired level of 0.8 for high-stakes testing. Training did not significantly improve assessment tool reliability. Although rater training may represent a way to improve reliability, further study is needed to determine effective training methods.
Trainee anaesthetist self-assessment using an entrustment scale in workplace-based assessment.
Castanelli D, Woods J, Chander A, Weller J Anaesth Intensive Care. 2024; 52(4):241-249.
PMID: 38649296 PMC: 11290023. DOI: 10.1177/0310057X241234676.
Faculty Perceptions of Frame of Reference Training to Improve Workplace-Based Assessment.
Kogan J, Conforti L, Holmboe E J Grad Med Educ. 2023; 15(1):81-91.
PMID: 36817545 PMC: 9934818. DOI: 10.4300/JGME-D-22-00287.1.
Interrater reliability in the assessment of physiotherapy students.
Gittinger F, Lemos M, Neumann J, Forster J, Dohmen D, Berke B BMC Med Educ. 2022; 22(1):186.
PMID: 35296313 PMC: 8928589. DOI: 10.1186/s12909-022-03231-y.
How can surgical skills in laparoscopic colon surgery be objectively assessed?-a scoping review.
Haug T, Worm Orntoft M, Miskovic D, Iversen L, Johnsen S, Madsen A Surg Endosc. 2021; 36(3):1761-1774.
PMID: 34873653 PMC: 8847271. DOI: 10.1007/s00464-021-08914-z.
Nayahangan L, Svendsen M, Bodtger U, Rahman N, Maskell N, Sidhu J J Thorac Dis. 2021; 13(7):3998-4007.
PMID: 34422330 PMC: 8339737. DOI: 10.21037/jtd-20-3560.