Leveraging Analytic Gradients in Provably Safe Reinforcement Learning

Tim Walter; Hannah Markgraf; Jonathan K\"ulz; and Matthias Althoff

arXiv:2506.01665·cs.LG·May 8, 2026

Leveraging Analytic Gradients in Provably Safe Reinforcement Learning

Tim Walter, Hannah Markgraf, Jonathan K\"ulz, and Matthias Althoff

PDF

1 Repo

TL;DR

This paper introduces the first effective safety safeguard for analytic gradient-based reinforcement learning, enhancing safety guarantees in autonomous robot training without sacrificing performance.

Contribution

It develops and integrates a novel safeguard into analytic gradient RL, addressing a key gap in safety guarantees for this learning paradigm.

Findings

01

Safeguarded training maintains performance levels.

02

Different safeguards influence learning outcomes.

03

Safeguards effectively improve safety during training.

Abstract

The deployment of autonomous robots in safety-critical applications requires safety guarantees. Provably safe reinforcement learning is an active field of research that aims to provide such guarantees using safeguards. These safeguards should be integrated during training to reduce the sim-to-real gap. While there are several approaches for safeguarding sampling-based reinforcement learning, analytic gradient-based reinforcement learning often achieves superior performance from fewer environment interactions. However, there is no safeguarding approach for this learning paradigm yet. Our work addresses this gap by developing the first effective safeguard for analytic gradient-based reinforcement learning. We analyse existing, differentiable safeguards, adapt them through modified mappings and gradient formulations, and integrate them into a state-of-the-art learning algorithm and a…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

null
github

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.