The “nonoverlapping intervals” and “reliable difference” approaches for assessing difference scores are compared and shown to be consistent alternatives when the proper z is used to construct confidence intervals. Formulas for computing the probabilities of correct interpretation (power), overinterpretation, misinterpretation, and underinterpretation for four popular confidence interval approaches and the reliable difference approach are presented. The probability formulas show that the intuitive inference concerning the statistical significance level of nonoverlapping intervals is incorrect. The limitations of the nonoverlapping intervals approach in applied situations are discussed. It appears that in most situations the reliable difference is the easiest to apply.
No takes yet. Share an insight, caveat, or question.
Charter et al. (2000) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: