VITS : Variational Inference Thomson Sampling for contextual bandits

07/19/2023
by   Pierre Clavier, et al.
0

In this paper, we introduce and analyze a variant of the Thompson sampling (TS) algorithm for contextual bandits. At each round, traditional TS requires samples from the current posterior distribution, which is usually intractable. To circumvent this issue, approximate inference techniques can be used and provide samples with distribution close to the posteriors. However, current approximate techniques yield to either poor estimation (Laplace approximation) or can be computationally expensive (MCMC methods, Ensemble sampling...). In this paper, we propose a new algorithm, Varational Inference Thompson sampling VITS, based on Gaussian Variational Inference. This scheme provides powerful posterior approximations which are easy to sample from, and is computationally efficient, making it an ideal choice for TS. In addition, we show that VITS achieves a sub-linear regret bound of the same order in the dimension and number of round as traditional TS for linear contextual bandit. Finally, we demonstrate experimentally the effectiveness of VITS on both synthetic and real world datasets.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/22/2022

Langevin Monte Carlo for Contextual Bandits

We study the efficiency of Thompson sampling for contextual bandits. Exi...
research
11/11/2022

Thompson Sampling for High-Dimensional Sparse Linear Contextual Bandits

We consider the stochastic linear contextual bandit problem with high-di...
research
10/30/2019

Thompson Sampling via Local Uncertainty

Thompson sampling is an efficient algorithm for sequential decision maki...
research
06/30/2023

Thompson sampling for improved exploration in GFlowNets

Generative flow networks (GFlowNets) are amortized variational inference...
research
05/07/2021

Laplace Matching for fast Approximate Inference in Generalized Linear Models

Bayesian inference in generalized linear models (GLMs), i.e. Gaussian re...
research
03/02/2022

An Analysis of Ensemble Sampling

Ensemble sampling serves as a practical approximation to Thompson sampli...
research
09/17/2019

Refined α-Divergence Variational Inference via Rejection Sampling

We present an approximate inference method, based on a synergistic combi...

Please sign up or login with your details

Forgot password? Click here to reset