Balancing Rational and Other-Regarding Preferences in Cooperative-Competitive Environments

02/24/2021
by   Dmitry Ivanov, et al.
8

Recent reinforcement learning studies extensively explore the interplay between cooperative and competitive behaviour in mixed environments. Unlike cooperative environments where agents strive towards a common goal, mixed environments are notorious for the conflicts of selfish and social interests. As a consequence, purely rational agents often struggle to achieve and maintain cooperation. A prevalent approach to induce cooperative behaviour is to assign additional rewards based on other agents' well-being. However, this approach suffers from the issue of multi-agent credit assignment, which can hinder performance. This issue is efficiently alleviated in cooperative setting with such state-of-the-art algorithms as QMIX and COMA. Still, when applied to mixed environments, these algorithms may result in unfair allocation of rewards. We propose BAROCCO, an extension of these algorithms capable to balance individual and social incentives. The mechanism behind BAROCCO is to train two distinct but interwoven components that jointly affect each agent's decisions. Our meta-algorithm is compatible with both Q-learning and Actor-Critic frameworks. We experimentally confirm the advantages over the existing methods and explore the behavioural aspects of BAROCCO in two mixed multi-agent setups.

READ FULL TEXT

page 6

page 8

page 9

page 10

page 11

research
06/07/2017

Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments

We explore deep reinforcement learning methods for multi-agent domains. ...
research
11/27/2021

Normative Disagreement as a Challenge for Cooperative AI

Cooperation in settings where agents have both common and conflicting in...
research
08/27/2014

Definition and properties to assess multi-agent environments as social intelligence tests

Social intelligence in natural and artificial systems is usually measure...
research
06/05/2019

Escaping the State of Nature: A Hobbesian Approach to Cooperation in Multi-agent Reinforcement Learning

Cooperation is a phenomenon that has been widely studied across many dif...
research
08/01/2018

Cooperative Group Optimization with Ants (CGO-AS): Leverage Optimization with Mixed Individual and Social Learning

We present CGO-AS, a generalized Ant System (AS) implemented in the fram...
research
05/10/2023

Learning Optimal "Pigovian Tax" in Sequential Social Dilemmas

In multi-agent reinforcement learning, each agent acts to maximize its i...
research
03/07/2022

Efficient Cooperation Strategy Generation in Multi-Agent Video Games via Hypergraph Neural Network

The performance of deep reinforcement learning (DRL) in single-agent vid...

Please sign up or login with your details

Forgot password? Click here to reset