A Lyapunov-Guided Post-Action Shield for Stability-Aware Deep Reinforcement Learning

nbnfi-fe20260907123261.pdf
Hyväksytty kirjoittajan käsikirjoitus - 2.92 MB
Kahil, H., Välisuo, P., & Elmusrati, M. (2026). A Lyapunov-Guided Post-Action Shield for Stability-Aware Deep Reinforcement Learning. In 2026 12th International Conference on Control, Decision and Information Technologies (CoDIT), 1649-1654. IEEE. https://doi.org/10.1109/CoDIT70676.2026.11630899
© 2026 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works.
Lataukset70

Kuvaus

This paper proposes a Lyapunov-guided post-action shielding mechanism for deep reinforcement learning (DRL) controllers under bounded actuation disturbances. In addition, an energy-safety requirement is formulated as a one-step energy threshold constraint that keeps the predicted next-state energy proxy within a prescribed limit, using a bounded-disturbance worst-case check. Simulation results show that the proposed mechanism substantially reduces constraint violations under actuation noise while preserving the nominal policy behavior whenever possible.

Emojulkaisu

2026 12th International Conference on Control, Decision and Information Technologies (CoDIT)

ISBN

979-8-3195-2077-7

ISSN

2576-3555
2576-3547

Aihealue

Kausijulkaisu

International conference on control, decision and information technologies

OKM-julkaisutyyppi

A4 Vertaisarvioitu artikkeli konferenssijulkaisussa