Learning Mixtures of Markov Chains and MDPs

11/17/2022
by   Chinmaya Kausik, et al.
0

We present an algorithm for use in learning mixtures of both Markov chains (MCs) and Markov decision processes (offline latent MDPs) from trajectories, with roots dating back to the work of Vempala and Wang. This amounts to handling Markov chains with optional control input. The method is modular in nature and amounts to (1) a subspace estimation step, (2) spectral clustering of trajectories, and (3) a few iterations of the EM algorithm. We provide end-to-end performance guarantees where we only explicitly require the number of trajectories to be linear in states and the trajectory length to be linear in mixing time. Experimental results suggest it outperforms both EM (95.4 average) and a previous method by Gupta et al. (54.1 accuracy on an 8x8 gridworld.

READ FULL TEXT
research
11/29/2021

On some mixing properties of copula-based Markov chains

This paper brings some insights of ψ'-mixing, ψ^*-mixing and ψ-mixing fo...
research
02/09/2023

Learning Mixtures of Markov Chains with Quality Guarantees

A large number of modern applications ranging from listening songs onlin...
research
04/19/2020

Faster Algorithms for Quantitative Analysis of Markov Chains and Markov Decision Processes with Small Treewidth

Discrete-time Markov Chains (MCs) and Markov Decision Processes (MDPs) a...
research
04/15/2021

Stochastic Processes with Expected Stopping Time

Markov chains are the de facto finite-state model for stochastic dynamic...
research
06/05/2021

Navigating to the Best Policy in Markov Decision Processes

We investigate the classical active pure exploration problem in Markov D...
research
07/29/2015

A Gauss-Newton Method for Markov Decision Processes

Approximate Newton methods are a standard optimization tool which aim to...
research
05/06/2018

Velocity formulae between entropy and hitting time for Markov chains

In the absence of acceleration, the velocity formula gives "distance tra...

Please sign up or login with your details

Forgot password? Click here to reset