Khan, Abdul Manan (2026) Stability-guaranteed reinforcement learning for Model Predictive Control gain tuning with transfer learning in Multi-UAV Systems. In: The 2nd International Conference on Smart Mobility and Logistics Ecosystems (SMILE 2026), 8-11 February 2026, Dhahran, Kingdom of Saudi Arabia,.
Preview |
PDF
KhanAM_Stability guaranteed reinforcement learning_VoR.pdf - Published Version Available under License Creative Commons Attribution Non-commercial No Derivatives. Download (486kB) | Preview |
Abstract
Tuning Model Predictive Control (MPC) weight matrices remains a tedious process that demands considerable expertise. A method is presented that enables a reinforcement learning agent to adjust gains online while enforcing Lyapunov-based bounds guaranteeing that every proposed gain lies inside a provably stable region. The resulting adaptive controller maintains stability regardless of policy network outputs. Across four Unmanned Aerial Vehicle (UAV) platforms spanning a 200-fold mass range (27 g to 5.5 kg), tracking improvements of 22–27% are observed on an aggressive 3D figure-8 trajectory spanning ±4.0 m on all axes (X, Y, and Z). Position root mean square error (RMSE) is reduced from 0.45–0.55 m to 0.33–0.43 m across all platforms, with variance reductions of 28–33%. No stability violations occurred throughout 60 evaluation trials. A projection operator clips out-of-bounds gains before they reach the MPC solver, acting as a hard safety layer. Sequential transfer learning reduces per-platform training by 75%. These findings demonstrate that formal stability constraints and learning-based adaptation can coexist
effectively.
| Item Type: | Conference or Workshop Item (Paper) |
|---|---|
| ISSN: | 2352-1465 |
| Page Range: | pp. 948-955 |
| Identifier: | 10.1016/j.trpro.2026.04.077 |
| Keywords: | MPC ; Reinforcement Learning ; Lyapunov Stability ; UAV Control ; Transfer Learning |
| Subjects: | Computing |
| Date Deposited: | 29 Sep 2026 |
| Dates: | Date Publication status 1 January 2026 Accepted 11 February 2026 Presented 2026 Published |
| School, department or research centre: | School of Computing and Engineering |
| Keywords: | MPC ; Reinforcement Learning ; Lyapunov Stability ; UAV Control ; Transfer Learning |
| URI: | https://repository.uwl.ac.uk/id/eprint/15445 | Sustainable Development Goals: | Goal 9: Industry, Innovation, and Infrastructure |
Downloads
Downloads per month over past year
Actions (admin access)
![]() |
Lists
Lists