Interpretable Reinforcement Learning with Multilevel Subgoal Discovery

Alexander Demin; Denis Ponomaryov

arXiv:2202.07414·cs.AI·February 16, 2022

Interpretable Reinforcement Learning with Multilevel Subgoal Discovery

Alexander Demin, Denis Ponomaryov

PDF

Open Access

TL;DR

This paper introduces an interpretable reinforcement learning model that discovers hierarchical subgoals in discrete environments without requiring reward functions, using probabilistic rules and state descriptions for efficient policy learning.

Contribution

It presents a novel RL framework that learns environment rules and subgoal hierarchies in an interpretable manner without reward signals.

Findings

01

Supports hierarchical subgoal discovery

02

Enables interpretable policy learning

03

Improves efficiency through state-based subgoals

Abstract

We propose a novel Reinforcement Learning model for discrete environments, which is inherently interpretable and supports the discovery of deep subgoal hierarchies. In the model, an agent learns information about environment in the form of probabilistic rules, while policies for (sub)goals are learned as combinations thereof. No reward function is required for learning; an agent only needs to be given a primary goal to achieve. Subgoals of a goal G from the hierarchy are computed as descriptions of states, which if previously achieved increase the total efficiency of the available policies for G. These state descriptions are introduced as new sensor predicates into the rule language of the agent, which allows for sensing important intermediate states and for updating environment rules and policies accordingly.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsNeural Networks and Applications