General non-linear Bellman equations

07/08/2019
by   Hado van Hasselt, et al.
5

We consider a general class of non-linear Bellman equations. These open up a design space of algorithms that have interesting properties, which has two potential advantages. First, we can perhaps better model natural phenomena. For instance, hyperbolic discounting has been proposed as a mathematical model that matches human and animal data well, and can therefore be used to explain preference orderings. We present a different mathematical model that matches the same data, but that makes very different predictions under other circumstances. Second, the larger design space can perhaps lead to algorithms that perform better, similar to how discount factors are often used in practice even when the true objective is undiscounted. We show that many of the resulting Bellman operators still converge to a fixed point, and therefore that the resulting algorithms are reasonable and inherit many beneficial properties of their linear counterparts.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/07/2019

Well-posedness study of a non-linear hyperbolic-parabolic coupled system applied to image speckle reduction

In this article, we consider a non-linear hyperbolic-parabolic coupled s...
research
11/29/2020

Applications of the Backus-Gilbert method to linear and some non linear equations

We investigate the use of a functional analytical version of the Backus-...
research
06/15/2017

Computational Anatomy in Theano

To model deformation of anatomical shapes, non-linear statistics are req...
research
04/09/2022

Generalised Mathematical Formulations for Non-Linear Optimized Scheduling

In practice, most of the optimization problems are non-linear requiring ...
research
03/01/2021

Hyperbolic relaxation technique for solving the dispersive Serre Equations with topography

The objective of this note is to propose a relaxation technique that acc...
research
02/12/2023

Statistical Complexity and Optimal Algorithms for Non-linear Ridge Bandits

We consider the sequential decision-making problem where the mean outcom...
research
05/06/2021

Concurrency Theorems for Non-linear Rewriting Theories

Sesqui-pushout (SqPO) rewriting along non-linear rules and for monic mat...

Please sign up or login with your details

Forgot password? Click here to reset