PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 18, 20260 citationsOpen Access

Model Selection Criteria CV, UBR, and GCV for a Mixed Truncated Spline–Gaussian Kernel Estimator in Health Modeling

IBI Nyoman BudiantaraNCNur ChamidahADAndrea Tri Rian Dani

Key Points

  • This research aims to assess the effectiveness of three smoothing parameter selection criteria for a mixed estimator in health modeling.
  • Investigated a Mixed Truncated Spline–Gaussian Kernel Estimator.
  • Compared Cross-Validation (CV), Unbiased Risk (UBR), and Generalized Cross-Validation (GCV) for selecting smoothing parameters.
  • Analyzed heart disease risk factor data characterized by complex relationships.
  • Models with three knot points showed better performance than alternatives.
  • Lower Mean Squared Error (MSE) and Mean Absolute Percentage Error (MAPE) were observed.
  • The Generalized Cross-Validation (GCV) criterion provided the most accurate and stable model.

Abstract

Nonparametric regression modeling has commonly applied a single estimator to all predictor variables. Although this approach is straightforward, it can be overly restrictive because predictors often exhibit heterogeneous relationships with the response variable, including linear trends, smooth nonlinear patterns, abrupt changes, or localized variations. Using a uniform estimator may therefore limit model flexibility and reduce predictive accuracy. To overcome this limitation, this study investigates a Mixed Truncated Spline–Gaussian Kernel Estimator, which allows each predictor to be modeled using the estimation technique most appropriate to its underlying data structure. The main objective of this research is to compare the performance of three smoothing parameter selection criteria, namely Cross-Validation (CV), Unbiased Risk (UBR), and Generalized Cross-Validation (GCV). These criteria are employed to determine optimal smoothing parameters, including the number and locations of knots in the truncated spline component and the bandwidth in the Gaussian kernel component. The empirical analysis is conducted using health-related data on heart disease risk factors, a domain characterized by complex and potentially nonlinear relationships. The results indicate that models incorporating three knot points consistently outperform alternative specifications. This superior performance is reflected in lower values of Mean Squared Error (MSE) and Mean Absolute Percentage Error (MAPE), as well as higher coefficients of determination (R²). Among the selection criteria examined, GCV yields the most accurate and stable model, outperforming both CV and UBR. From a methodological perspective, this study contributes to nonparametric regression by providing a systematic evaluation of smoothing parameter selection within a mixed estimator framework. From an applied standpoint, the proposed approach enhances the modeling of heart disease risk factors by offering greater flexibility and precision. Furthermore, the findings support Sustainable Development Goal (SDG) 3: Good Health and Well-Being by promoting robust, data-driven methods for evidence-based health policy formulation.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Budiantara et al. (2026) studied this question.

synapsesocial.com/papers/69e320fd40886becb65401f2https://doi.org/10.19139/soic-2310-5070-3472
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Penalized Spline Estimator for Semiparametric Binary Logistic Regression Model with Application to Coronary Heart Disease Risk Factors2026
  2. 2Variable selection in functional linear Cox model2026
  3. 3Stunting Prevalence Modeling Using Nonparametric Regression of Quadratic Splines2024 · 3 citations
  4. 4Penalized Spline Semiparametric Logistic Regression for Modelling Coronary Heart Disease Risk Based on Demographic and Lifestyle Factors2026
  5. 5Using mixed kernel support vector machine to improve the predictive accuracy of genome selection2024 · 6 citations