We present an estimate of the kappa-coefficient of agreement between two methods of rating based on matched pairs of binary responses and show that the estimate depends on the common intraclass correlation coefficient between the pairs. Via Monte Carlo simulation, we investigate power of the test of significance on kappa, and the large sample bias and variance of its maximum likelihood estimator.
No takes yet. Share an insight, caveat, or question.
Shoukri et al. (1995) studied this question.
Synapse has enriched 3 closely related papers on similar clinical questions. Consider them for comparative context: