PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 14, 2026Aerospace0 citationsOpen Access

Hierarchical Adaptive PID Tuning for Agile Flight: A Safety-Constrained Reinforcement Learning Approach

View Full Paper
ZTZhong TianSHSen HuHFHao Fu

Key Points

  • This research aims to enhance control performance of multirotor UAVs during aggressive maneuvers using adaptive PID tuning.
  • Developed a hierarchical safety-constrained reinforcement learning framework for adaptive PID tuning.
  • Used proximal policy optimization for online adaptive gain scheduling.
  • Applied linear matrix inequality constraints to ensure robust parameter boundaries.
  • Reduced overshoot by 18.5% in high-speed step responses compared to traditional PID controllers.
  • Improved overall mean RMSE by 15.0% across 100 randomized mixed-trajectory trials.
  • Achieved up to 40.9% improvement in highly dynamic scenarios for trajectory tracking accuracy.

Abstract

Multirotor unmanned aerial vehicles (UAVs) suffer from significant control performance degradation during aggressive maneuvers, primarily due to aerodynamic nonlinearities and coupling effects. Conventional fixed-gain PID controllers struggle to simultaneously satisfy performance and robustness requirements across the wide flight envelope. To address this challenge, this paper presents a novel hierarchical safety-constrained reinforcement learning (RL) framework for adaptive PID tuning: the inner loop employs fixed gains, the outer loop leverages proximal policy optimization (PPO) for online adaptive gain scheduling, and linear matrix inequality (LMI) constraints delineate robust parameter boundaries for the adaptive exploration. Importantly, the LMI feasibility strictly guarantees theoretical stability for the fixed inner-loop parameters at the linearization vertices within a linear parameter-varying (LPV) framework. Concurrently, the online outer-loop RL stage is protected by safety boundaries and a Lagrangian penalty mechanism acting as an effective engineering safeguard rather than a rigorous global stability proof. Comprehensive high-fidelity simulation benchmarks demonstrate that, compared with a baseline fixed-gain PID controller, the proposed framework reduces overshoot by 18.5% in high-speed step responses and improves the overall mean RMSE by 15.0% across 100 randomized mixed-trajectory trials (with improvements of up to 40.9% in highly dynamic scenarios), yielding consistent gains in trajectory tracking accuracy and disturbance rejection despite uncertain model variations. By seamlessly blending control-theoretic hard constraints with RL-based soft-parameter tuning, the proposed architecture offers a safe and highly adaptive solution for large-envelope flight control, demonstrating strong engineering relevance.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Tian et al. (2026) studied this question.

synapsesocial.com/papers/6a0567bca550a87e60a1fdb6https://doi.org/10.3390/aerospace13050446
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Reinforcement Learning-Based PD Controller Gains Prediction for Quadrotor UAVs2025 · 6 citations
  2. 2Reinforcement Learning Based PID Parameter Tuning and Estimation for Multirotor UAVs2024 · 12 citations
  3. 3Performance Analysis of Q-Learning; DQN-Enhanced PID Controllers for Aircraft Trajectory Tracking Under Wind Disturbances2026
  4. 4AirPilot: Interpretable PPO-based DRL Auto-Tuned Nonlinear PID Drone Controller for Robust Autonomous Flights2024 · 1 citations
  5. 5Robust fractional-order adaptive gain-scheduled control strategy for civil unmanned aerial vehicle with LPV models2025 · 1 citations