Adaptive Surrogate Gradients for Sequential Reinforcement Learning in Spiking Neural Networks

Korneel Van den Berghe; Stein Stroobants; Vijay Janapa Reddi; G.C.H.E. de Croon

arXiv:2510.24461·cs.AI·October 29, 2025

Adaptive Surrogate Gradients for Sequential Reinforcement Learning in Spiking Neural Networks

Korneel Van den Berghe, Stein Stroobants, Vijay Janapa Reddi, G.C.H.E. de Croon

PDF

1 Video

TL;DR

This paper introduces adaptive surrogate gradients and a guiding policy to improve training of spiking neural networks for reinforcement learning, achieving significant performance gains in real-world robotic control.

Contribution

It provides a systematic analysis of surrogate gradient slopes and proposes an adaptive slope schedule combined with a guiding policy for effective RL training in SNNs.

Findings

01

Shallower slopes increase gradient magnitude in deep layers but reduce alignment.

02

Shallow or scheduled slopes improve RL performance by 2.1x.

03

The proposed method outperforms prior techniques in drone control, achieving an average return of 400 points.

Abstract

Neuromorphic computing systems are set to revolutionize energy-constrained robotics by achieving orders-of-magnitude efficiency gains, while enabling native temporal processing. Spiking Neural Networks (SNNs) represent a promising algorithmic approach for these systems, yet their application to complex control tasks faces two critical challenges: (1) the non-differentiable nature of spiking neurons necessitates surrogate gradients with unclear optimization properties, and (2) the stateful dynamics of SNNs require training on sequences, which in reinforcement learning (RL) is hindered by limited sequence lengths during early training, preventing the network from bridging its warm-up period. We address these challenges by systematically analyzing surrogate gradient slope settings, showing that shallower slopes increase gradient magnitude in deeper layers but reduce alignment with true…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

Adaptive Surrogate Gradients for Sequential Reinforcement Learning in Spiking Neural Networks· slideslive