A test statistic is introduced which allows one to test the hypothesis of agreement of several judges on the ranking of items within each of two groups and between the two groups. The groups of judges may be unequal in size. A normal approximation for the test statistic is developed. The relationship to existing techniques given by Kendall, Friedman, Page, Spearman, and Lyerly is discussed. A generalization of the coefficient of concordance is presented and the extension of the method to multi-group problems is suggested.
No takes yet. Share an insight, caveat, or question.
Schucany et al. (1973) studied this question.
Synapse has enriched 2 closely related papers on similar clinical questions. Consider them for comparative context: