Online Contextual Decision-Making with a Smart Predict-then-Optimize Method

06/15/2022
by   Heyuan Liu, et al.
0

We study an online contextual decision-making problem with resource constraints. At each time period, the decision-maker first predicts a reward vector and resource consumption matrix based on a given context vector and then solves a downstream optimization problem to make a decision. The final goal of the decision-maker is to maximize the summation of the reward and the utility from resource consumption, while satisfying the resource constraints. We propose an algorithm that mixes a prediction step based on the "Smart Predict-then-Optimize (SPO)" method with a dual update step based on mirror descent. We prove regret bounds and demonstrate that the overall convergence rate of our method depends on the 𝒪(T^-1/2) convergence of online mirror descent as well as risk bounds of the surrogate loss function used to learn the prediction model. Our algorithm and regret bounds apply to a general convex feasible region for the resource constraints, including both hard and soft resource constraint cases, and they apply to a wide class of prediction models in contrast to the traditional settings of linear contextual models or finite policy spaces. We also conduct numerical experiments to empirically demonstrate the strength of our proposed SPO-type methods, as compared to traditional prediction-error-only methods, on multi-dimensional knapsack and longest path instances.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
08/19/2021

Risk Bounds and Calibration for a Smart Predict-then-Optimize Method

The predict-then-optimize framework is fundamental in practical stochast...
research
04/20/2021

Joint Online Learning and Decision-making via Dual Mirror Descent

We consider an online revenue maximization problem over a finite time ho...
research
11/22/2021

A Surrogate Objective Framework for Prediction+Optimization with Soft Constraints

Prediction+optimization is a common real-world paradigm where we have to...
research
11/09/2022

A Note on Task-Aware Loss via Reweighing Prediction Loss by Decision-Regret

In this short technical note we propose a baseline for decision-aware le...
research
04/14/2022

Gradient boosting for convex cone predict and optimize problems

Many problems in engineering and statistics involve both predictive fore...
research
04/30/2023

Electricity Price Prediction for Energy Storage System Arbitrage: A Decision-focused Approach

Electricity price prediction plays a vital role in energy storage system...
research
11/05/2020

Fast Rates for Contextual Linear Optimization

Incorporating side observations of predictive features can help reduce u...

Please sign up or login with your details

Forgot password? Click here to reset