Large Scale Learning of Agent Rationality in Two-Player Zero-Sum Games

by   Chun Kai Ling, et al.

With the recent advances in solving large, zero-sum extensive form games, there is a growing interest in the inverse problem of inferring underlying game parameters given only access to agent actions. Although a recent work provides a powerful differentiable end-to-end learning frameworks which embed a game solver within a deep-learning framework, allowing unknown game parameters to be learned via backpropagation, this framework faces significant limitations when applied to boundedly rational human agents and large scale problems, leading to poor practicality. In this paper, we address these limitations and propose a framework that is applicable for more practical settings. First, seeking to learn the rationality of human agents in complex two-player zero-sum games, we draw upon well-known ideas in decision theory to obtain a concise and interpretable agent behavior model, and derive solvers and gradients for end-to-end learning. Second, to scale up to large, real-world scenarios, we propose an efficient first-order primal-dual method which exploits the structure of extensive-form games, yielding significantly faster computation for both game solving and gradient computation. When tested on randomly generated games, we report speedups of orders of magnitude over previous approaches. We also demonstrate the effectiveness of our model on both real-world one-player settings and synthetic data.


page 1

page 2

page 3

page 4


What game are we playing? End-to-end learning in normal and extensive form games

Although recent work in AI has made great progress in solving large, zer...

LP Formulations of Two-Player Zero-Sum Stochastic Bayesian games

This paper studies two-player zero-sum stochastic Bayesian games where e...

A Generic Multi-Player Transformation Algorithm for Solving Large-Scale Zero-Sum Extensive-Form Adversarial Team Games

Many recent practical and theoretical breakthroughs focus on adversarial...

Convergence analysis and acceleration of the smoothing methods for solving extensive-form games

The extensive-form game has been studied considerably in recent years. I...

Cumulative Games: Who is the current player?

Combinatorial Game Theory (CGT) is a branch of game theory that has deve...

Solution of Two-Player Zero-Sum Game by Successive Relaxation

We consider the problem of two-player zero-sum game. In this setting, th...

Solving Atari Games Using Fractals And Entropy

In this paper, we introduce a novel MCTS based approach that is derived ...

Please sign up or login with your details

Forgot password? Click here to reset