Markov Decision Process with an External Temporal Process

05/25/2023

∙

Most reinforcement learning algorithms treat the context under which they operate as a stationary, isolated and undisturbed environment. However, in the real world, the environment is constantly changing due to a variety of external influences. To address this problem, we study Markov Decision Processes (MDP) under the influence of an external temporal process. We formalize this notion and discuss conditions under which the problem becomes tractable with suitable solutions. We propose a policy iteration algorithm to solve this problem and theoretically analyze its performance.

READ FULL TEXT

Markov Decision Process with an External Temporal Process

Sign in with Google

Consider DeepAI Pro