We create a hedonic price model for house prices for six geographical submarkets in the Netherlands. Our model is based on a recent data‐mining technique called boosting. Boosting is an ensemble technique that combines multiple models, in our case decision trees, into a combined prediction. Boosting enables capturing of complex nonlinear relationships and interaction effects between input variables. We report mean relative errors and mean absolute error for all regions and compare our models with a standard linear regression approach. Our model improves prediction performance by up to 39% compared with linear regression and by up to 20% compared with a log‐linear regression model. Next, we interpret the boosted models: we determine the most influential characteristics and graphically depict the relationship between the most important input variables and the house price. We find the size of the house to be the most important input for all but one region, and find some interesting nonlinear relationships between inputs and price. Finally, we construct hedonic price indices and compare these with the mean and median index and find that these indices differ notably in the urban regions of Amsterdam and Rotterdam.
No takes yet. Share an insight, caveat, or question.
Kagie et al. (2007) studied this question.
Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context: