Proximal Policy Optimization Learning based Control of Congested Freeway Traffic

04/12/2022
by   Shurong Mo, et al.
6

This study proposes a delay-compensated feedback controller based on proximal policy optimization (PPO) reinforcement learning to stabilize traffic flow in the congested regime by manipulating the time-gap of adaptive cruise control-equipped (ACC-equipped) vehicles.The traffic dynamics on a freeway segment are governed by an Aw-Rascle-Zhang (ARZ) model, consisting of 2× 2 nonlinear first-order partial differential equations (PDEs).Inspired by the backstepping delay compensator [18] but different from whose complex segmented control scheme, the PPO control is composed of three feedbacks, namely the current traffic flow velocity, the current traffic flow density and previous one step control input. The control gains for the three feedbacks are learned from the interaction between the PPO and the numerical simulator of the traffic system without knowing the system dynamics. Numerical simulation experiments are designed to compare the Lyapunov control, the backstepping control and the PPO control. The results show that for a delay-free system, the PPO control has faster convergence rate and less control effort than the Lyapunov control. For a traffic system with input delay, the performance of the PPO controller is comparable to that of the Backstepping controller, even for the situation that the delay value does not match. However, the PPO is robust to parameter perturbations, while the Backstepping controller cannot stabilize a system where one of the parameters is disturbed by Gaussian noise.

READ FULL TEXT

page 1

page 6

page 7

page 8

research
05/28/2021

Feedback Linearization for Quadrotors with a Learned Acceleration Error Model

This paper enhances the feedback linearization controller for multirotor...
research
01/17/2019

Multiclass Information Flow Propagation Control under Vehicle-to-Vehicle Communication Environments

Most existing models for information flow propagation in a vehicle-to-ve...
research
12/27/2021

Intelligent Traffic Light via Policy-based Deep Reinforcement Learning

Intelligent traffic lights in smart cities can optimally reduce traffic ...
research
03/26/2020

Properties of the LWR model with time delay

In this article, we investigate theoretical and numerical properties of ...
research
01/18/2023

Hierarchical Reinforcement Learning Based Traffic Steering in Multi-RAT 5G Deployments

In 5G non-standalone mode, an intelligent traffic steering mechanism can...
research
08/28/2019

Networked Control of Nonlinear Systems under Partial Observation Using Continuous Deep Q-Learning

In this paper, we propose a design of a model-free networked controller ...
research
05/16/2022

Towards on-sky adaptive optics control using reinforcement learning

The direct imaging of potentially habitable Exoplanets is one prime scie...

Please sign up or login with your details

Forgot password? Click here to reset