Learning Markov models via low-rank optimization

06/28/2019
by   Ziwei Zhu, et al.
0

Modeling unknown systems from data is a precursor of system optimization and sequential decision making. In this paper, we focus on learning a Markov model from a single trajectory of states. Suppose that the transition model has a small rank despite of a large state space, meaning that the system admits a low-dimensional latent structure. We show that one can estimate the full transition model accurately using a trajectory of length that is proportional to the total number of states. We propose two maximum likelihood estimation methods: a convex approach with nuclear-norm regularization and a nonconvex approach with rank constraint. We show that both estimators enjoy optimal statistical rates in terms of the Kullback-Leiber divergence and the ℓ_2 error. For computing the nonconvex estimator, we develop a novel DC (difference of convex function) programming algorithm that starts with the convex M-estimator and then successively refines the solution till convergence. Empirical experiments demonstrate consistent superiority of the nonconvex estimator over the convex one.

READ FULL TEXT

page 25

page 26

research
04/03/2018

Estimation of Markov Chain via Rank-constrained Likelihood

This paper studies the recovery and state compression of low-rank Markov...
research
04/19/2019

Square-root nuclear norm penalized estimator for panel data models with approximately low-rank unobserved heterogeneity

This paper considers a nuclear norm penalized estimator for panel data m...
research
04/06/2020

Low-Rank Matrix Estimation From Rank-One Projections by Unlifted Convex Optimization

We study an estimator with a convex formulation for recovery of low-rank...
research
05/06/2022

Low-rank Tensor Learning with Nonconvex Overlapped Nuclear Norm Regularization

Nonconvex regularization has been popularly used in low-rank matrix lear...
research
09/22/2022

Maximum likelihood estimation for nonembeddable Markov chains when the cycle length is shorter than the data observation interval

Time-homogeneous Markov chains are often used as disease progression mod...
research
10/14/2018

Adaptive Low-Nonnegative-Rank Approximation for State Aggregation of Markov Chains

This paper develops a low-nonnegative-rank approximation method to ident...
research
09/27/2019

Identifying Low-Dimensional Structures in Markov Chains: A Nonnegative Matrix Factorization Approach

A variety of queries about stochastic systems boil down to study of Mark...

Please sign up or login with your details

Forgot password? Click here to reset