The ANOVA, Welch, and Brown and Forsyth tests for mean equality were compared using Monte Carlo methods. The tests’ rates of Type I error and power were examined when populations were non-normal, variances were heterogeneous, and group sizes were unequal. The ANOVA F test was most affected by the assumption violations. The test proposed by Brown and Forsyth appeared, on the average, to be the “best” test statistic for testing an omnibus hypothesis of mean equality.
No takes yet. Share an insight, caveat, or question.
Clinch et al. (1982) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: