Behavior-based Neuroevolutionary Training in Reinforcement Learning

J\"org Stork; Martin Zaefferer; Nils Eisler; Patrick Tichelmann,; Thomas Bartz-Beielstein; A. E. Eiben

arXiv:2105.07960·cs.NE·May 18, 2021

Behavior-based Neuroevolutionary Training in Reinforcement Learning

J\"org Stork, Martin Zaefferer, Nils Eisler, Patrick Tichelmann,, Thomas Bartz-Beielstein, A. E. Eiben

PDF

1 Repo

TL;DR

This paper introduces a hybrid neuroevolutionary and value-based reinforcement learning algorithm that improves sample efficiency and learning speed by exploiting stored experiences and behavior modeling.

Contribution

It presents a novel hybrid approach combining neuroevolution with value-based RL, utilizing behavior-based loss functions and directed search in behavior space.

Findings

01

Enhanced sample efficiency over traditional evolutionary methods

02

Faster learning speed demonstrated on benchmarks and real-world problem

03

Effective behavior modeling improves policy optimization

Abstract

In addition to their undisputed success in solving classical optimization problems, neuroevolutionary and population-based algorithms have become an alternative to standard reinforcement learning methods. However, evolutionary methods often lack the sample efficiency of standard value-based methods that leverage gathered state and value experience. If reinforcement learning for real-world problems with significant resource cost is considered, sample efficiency is essential. The enhancement of evolutionary algorithms with experience exploiting methods is thus desired and promises valuable insights. This work presents a hybrid algorithm that combines topology-changing neuroevolutionary optimization with value-based reinforcement learning. We illustrate how the behavior of policies can be used to create distance and loss functions, which benefit from stored experiences and calculated state…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

jstork/BNET-GECCO21
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.