A reading difficulty measure can be described as a function or model that maps a text to a numerical value corresponding to a difficulty or grade level. We describe a measure of readability that uses a combination of lexical features and grammatical features that are derived from subtrees of syntactic parses. We also tested statistical models for nominal, ordinal, and interval scales of measurement. The results indicate that a model for ordinal regression, such as the proportional odds model, using a combination of grammatical and lexical features is most effective at predicting reading difficulty.
No takes yet. Share an insight, caveat, or question.
Heilman et al. (2008) studied this question.
Synapse has enriched 2 closely related papers on similar clinical questions. Consider them for comparative context: