This article discusses and demonstrates combining scores from multiple‐choice (MC) and constructed‐response (CR) items to create a common scale using item response theory methodology. Two specific issues addressed are (a) whether MC and CR items can be calibrated together and (b) whether simultaneous calibration of the two item types leads to loss of information. Procedures are discussed and empirical results are provided using a set of tests in the areas of reading, language, mathematics, and science in three grades.
No takes yet. Share an insight, caveat, or question.
Ercikan et al. (1998) studied this question.
Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context: