General non-linear Bellman equations

by   Hado van Hasselt, et al.

We consider a general class of non-linear Bellman equations. These open up a design space of algorithms that have interesting properties, which has two potential advantages. First, we can perhaps better model natural phenomena. For instance, hyperbolic discounting has been proposed as a mathematical model that matches human and animal data well, and can therefore be used to explain preference orderings. We present a different mathematical model that matches the same data, but that makes very different predictions under other circumstances. Second, the larger design space can perhaps lead to algorithms that perform better, similar to how discount factors are often used in practice even when the true objective is undiscounted. We show that many of the resulting Bellman operators still converge to a fixed point, and therefore that the resulting algorithms are reasonable and inherit many beneficial properties of their linear counterparts.


page 1

page 2

page 3

page 4


Well-posedness study of a non-linear hyperbolic-parabolic coupled system applied to image speckle reduction

In this article, we consider a non-linear hyperbolic-parabolic coupled s...

Applications of the Backus-Gilbert method to linear and some non linear equations

We investigate the use of a functional analytical version of the Backus-...

Computational Anatomy in Theano

To model deformation of anatomical shapes, non-linear statistics are req...

Generalised Mathematical Formulations for Non-Linear Optimized Scheduling

In practice, most of the optimization problems are non-linear requiring ...

Hyperbolic relaxation technique for solving the dispersive Serre Equations with topography

The objective of this note is to propose a relaxation technique that acc...

Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits

We consider the sequential decision-making problem where the mean outcom...

Concurrency Theorems for Non-linear Rewriting Theories

Sesqui-pushout (SqPO) rewriting along non-linear rules and for monic mat...

Please sign up or login with your details

Forgot password? Click here to reset