Key points are not available for this paper at this time.
OBJECTIVES: A study was conducted to determine the reliability of Agency for Healthcare Research in 2012, a second survey evaluating Harm Scale v.1.2 was sent to 13,280 managers at 102 organizations. RESULTS: Regardless of the version used, in 3 of 9 scenarios, fewer than 60% of respondents agreed on a single score. Interrater agreement increased for certain event scenarios with v.1.2 but decreased for other scenarios. Interrater reliability was moderate for both v.1.1 (k = 0.51) and v.1.2 (k = 0.47). Interrater agreement improved in v.1.2 when results were limited to more experienced raters but still remained in the moderate range (k = 0.58). CONCLUSIONS: AHRQ Common Format Harm Scale v.1.1 and v.1.2 both had moderate interrater reliability. Using Harm Scale v.1.1, respondents had difficulty distinguishing "injury limited to additional treatment" from "temporary harm," whereas, using Harm Scale v.1.2, respondents had difficulty distinguishing moderate harm from one of the adjacent levels-mild or severe harm. This study provides valuable data that can inform harm scale revision to improve the quality of aggregate safety data used to define and direct safety efforts.
Williams et al. (Tue,) studied this question.