PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 17, 2026Polymers for Advanced Technologies0 citations

Hybrid Gradient Boosting Decision Tree to Accurately Compute Polyethylene Glycol Polymer Density

View Full Paper
JLJianfeng LiGuangdong Polytechnic of Science and TechnologySISoud Khalil IbrahimAl-Nisour University CollegeRJRamdevsinh JhalaMarwadi University

Key Points

  • The aim is to accurately predict the density of polyethylene glycol (PEG) using a machine learning model.
  • Developed a hybrid gradient boosting decision tree model
  • Optimized hyperparameters using four evolutionary algorithms
  • Compiled a dataset of 293 points from existing literature
  • Applied K-fold cross-validation to assess model performance
  • Temperature was identified as the strongest influencer on PEG density (relevance factor: -0.87)
  • The ES algorithm yielded the highest accuracy with R² of 0.996 and MSE of 1.149
  • BPI algorithm followed closely with an R² of 0.994 and MSE of 1.663
  • SADE showed the least effectiveness with R² of 0.989 and MSE of 1.939

Abstract

ABSTRACT Accurate prediction of polyethylene glycol (PEG) density is critical for its growing use as an eco‐friendly solvent in chemical separations and industrial processes, yet experimental measurements are time‐intensive and costly. This study addresses the challenge by developing a hybrid gradient boosting decision tree (GBDT) machine learning model to predict PEG density with high precision. The model's hyperparameters were optimized using four evolutionary algorithms including evolutionary strategies (ES), Bayesian probability improvement (BPI), batch Bayesian optimization (BBO), and self‐adaptive differential evolution (SADE) on a dataset of 293 points compiled from existing literature, with K ‐fold cross‐validation applied to prevent overfitting. Model performance was assessed via optimization runtime and metrics including R ‐squared ( R 2 ), mean squared error (MSE), and average absolute relative error percentage (AARE%). Key findings indicate that temperature has the strongest influence on PEG density (relevance factor: −0.87), followed by weaker correlations with pressure (0.39) and molecular weight (−0.17). The ES algorithm achieved the highest accuracy ( R 2 : 0.996, MSE: 1.149, AARE%: 0.078% on the test dataset), closely followed by BPI ( R 2 : 0.994, MSE: 1.663, AARE%: 0.092%), while SADE performed least effectively ( R 2 : 0.989, MSE: 1.939, AARE%: 0.091%) with the longest runtime (~2000 s). Sensitivity and SHAP analyses confirmed temperature's dominant role. These hybrid models offer a novel, computationally efficient alternative to experimental methods, advancing predictive modeling for PEG's physicochemical properties by integrating evolutionary optimization with GBDT, achieving superior accuracy compared to prior approaches.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Li et al. (2026) studied this question.

synapsesocial.com/papers/696b2696d2a12237a9349e28https://doi.org/10.1002/pat.70455
Ask AI
Helpful
Bookmark
Share
View Full Paper