科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Communications of the ACM2025-10-20· Reinforcement learning

Shields for Safe Reinforcement Learning

Bettina Könighofer, Roderick Bloem, Nils Jansen, Sebastian Junges, Stefan Pranger

原始摘要(英文原文)· Original abstract
Reinforcement learning (RL) is a prominent machine learning technique used to optimize an agent’s performance in potentially unknown environments. Despite its popularity and success, RL lacks safety guarantees, both during the learning phase and deployment. This paper reviews a runtime enforcement method called shielding that ensures provable safety for RL. We describe the underlying models, the types of guarantees that can be delivered, and the process of computing shields. Furthermore, we describe several techniques for integrating shields into RL, discuss the advantages and potential drawbacks of this integration, and highlight the current challenges in shielded learning.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Shields for Safe Reinforcement Learning — 科研速览 Science Skim