PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 17, 2026Engineering Applications of Artificial Intelligence1 citationsOpen Access

Integrating transfer learning into multi-agent reinforcement learning for energy-efficient adaptive control of heating systems in thermoforming

View Full Paper
IJIman JalilvandAEAmir M. Soufi EnayatiHHHadi Hosseinionari

Key Points

  • The aim is to improve energy efficiency and adaptability of heating systems in thermoforming using advanced learning techniques.
  • Developed a Fully Connected Neural Network-based Multi-Agent Proximal Policy Optimization architecture.
  • Implemented a multi-objective reward function to minimize control error and energy usage.
  • Established the Multi-Agent Transfer Reinforcement Learning framework for enhancing learning efficiency.
  • Reduced training time by 29.3% and energy consumption by 28.1% under high convective heat transfer conditions.
  • Achieved a temperature error reduction of 2.02 °C compared to baseline methods.
  • Maintained energy efficiency within 0.5% of nominal conditions while improving temperature dispersion by 47% under extreme variations.

Abstract

This study proposes a framework that integrates Transfer Learning (TL) into Multi-Agent Reinforcement Learning (MARL) for real-time, energy-efficient thermal control of the heating stage in thermoforming processes, while concurrently minimizing energy consumption and adapting to varying manufacturing conditions. To enable a scalable framework under these conditions, a Fully Connected Neural Network-based Multi-Agent Proximal Policy Optimization (FCNN-MAPPO) architecture is developed using a multi-objective reward function for concurrently minimizing control error, thermal energy consumption, and instability, while preserving temporal dynamics through state augmentation. The resulting Multi-Agent Transfer Reinforcement Learning (MATRL) framework combines direct parameter transfer for within-scenario learning with experience sharing for cross-scenario adaptation, enabling faster convergence and improved generalization. Results show that under high convective heat transfer conditions, MATRL can reduce training time by 29.3%, decrease energy consumption by 28.1%, and improve average error by 2.02 °C compared to baseline MARL (i.e., without TL). Under elevated ambient temperature conditions, energy usage and settling time were reduced by 42.6% and 41.7%, respectively. Under synchronous multi-parameter variations (extreme conditions of convective heat transfer, sheet conductivity, and ambient temperature), MATRL maintained energy efficiency within 0.5% of nominal conditions while reducing temperature dispersion by 47%, demonstrating robust multidimensional adaptability without retraining. Statistical validation across multiple runs with random seeds showed stable performance, with a coefficient of variation of 7.3% and no divergence.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Jalilvand et al. (2026) studied this question.

synapsesocial.com/papers/6a095c2c7880e6d24efe2283https://doi.org/10.1016/j.engappai.2026.114780
Ask AI
Helpful
Bookmark
Share
View Full Paper