Probably Approximately Correct MDP Learning and Control With Temporal   Logic Constraints

Jie Fu; Ufuk Topcu

arXiv:1404.7073·cs.SY·May 1, 2014·66 cites

Probably Approximately Correct MDP Learning and Control With Temporal Logic Constraints

Jie Fu, Ufuk Topcu

PDF

Open Access

TL;DR

This paper presents a PAC-MDP-based algorithm for synthesizing control policies that maximize the probability of satisfying temporal logic specifications in unknown stochastic environments, with guarantees on near-optimality and polynomial complexity.

Contribution

It introduces a model-based PAC-MDP approach for control synthesis under temporal logic constraints, ensuring near-optimal policies with polynomial sample complexity.

Findings

01

Algorithm achieves ε-approximate optimality with high probability

02

Policy construction scales polynomially with MDP and automaton size

03

Finitely terminating iterative policy updates

Abstract

We consider synthesis of control policies that maximize the probability of satisfying given temporal logic specifications in unknown, stochastic environments. We model the interaction between the system and its environment as a Markov decision process (MDP) with initially unknown transition probabilities. The solution we develop builds on the so-called model-based probably approximately correct Markov decision process (PAC-MDP) methodology. The algorithm attains an $ε$ -approximately optimal policy with probability $1 - δ$ using samples (i.e. observations), time and space that grow polynomially with the size of the MDP, the size of the automaton expressing the temporal logic specification, $\frac{1}{ε}$ , $\frac{1}{δ}$ and a finite time horizon. In this approach, the system maintains a model of the initially unknown MDP, and constructs a product MDP based on…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsFormal Methods in Verification · Machine Learning and Algorithms · Bayesian Modeling and Causal Inference