Shapley Q-value: A Local Reward Approach to Solve Global Reward Games

07/11/2019
by   Jianhong Wang, et al.
0

Cooperative game is a critical research area in multi-agent reinforcement learning (MARL). Global reward game is a subclass of cooperative games, where all agents aim to maximize cumulative global rewards. Credit assignment is an important problem studied in the global reward game. Most of works stand by the view of non-cooperative-game theoretical framework with the shared reward approach, i.e., each agent being assigned a shared global reward directly. This, however, may give each agent an inaccurate feedback on its contribution to the group. In this paper, we introduce a cooperative-game theoretical framework and extend it to the infinite-horizon case. We show that our proposed framework is a superset of the global reward game. Based on this framework, we propose a local reward approach called Shapley Q-value that can distribute the cumulative global rewards fairly, reflecting each agent's own contribution in contrast to the shared reward approach. Moreover, we derive an MARL algorithm called Shapley Q-value policy gradient (SQPG), using Shapley Q-value as critics. We evaluate SQPG on Cooperative Navigation, Prey-and-Predator and Traffic Junction, compared with MADDPG, COMA, Independent A2C and Independent DDPG. In the experiments, SQPG shows the better performance than the baselines. In addition, we also plot the Shapley Q-value and validate the property of fairly distributing the global rewards.

READ FULL TEXT
research
07/11/2019

Rethink Global Reward Game and Credit Assignment in Multi-agent Reinforcement Learning

Cooperative game is a critical research area in multi-agent reinforcemen...
research
10/22/2019

Teach Biped Robots to Walk via Gait Principles and Reinforcement Learning with Adversarial Critics

Controlling a biped robot to walk stably is a challenging task consideri...
research
03/05/2020

Reward Design in Cooperative Multi-agent Reinforcement Learning for Packet Routing

In cooperative multi-agent reinforcement learning (MARL), how to design ...
research
09/22/2021

Locality Matters: A Scalable Value Decomposition Approach for Cooperative Multi-Agent Reinforcement Learning

Cooperative multi-agent reinforcement learning (MARL) faces significant ...
research
01/30/2019

Coordinating the Crowd: Inducing Desirable Equilibria in Non-Cooperative Systems

Many real-world systems such as taxi systems, traffic networks and smart...
research
04/29/2021

Adapting to Reward Progressivity via Spectral Reinforcement Learning

In this paper we consider reinforcement learning tasks with progressive ...
research
09/25/2022

Cooperative Sensing and Heterogeneous Information Fusion in VCPS: A Multi-agent Deep Reinforcement Learning Approach

Cooperative sensing and heterogeneous information fusion are critical to...

Please sign up or login with your details

Forgot password? Click here to reset