Empirical Risk Minimization for Losses without Variance

09/07/2023
by   Guanhua Fang, et al.
0

This paper considers an empirical risk minimization problem under heavy-tailed settings, where data does not have finite variance, but only has p-th moment with p ∈ (1,2). Instead of using estimation procedure based on truncated observed data, we choose the optimizer by minimizing the risk value. Those risk values can be robustly estimated via using the remarkable Catoni's method (Catoni, 2012). Thanks to the structure of Catoni-type influence functions, we are able to establish excess risk upper bounds via using generalized generic chaining methods. Moreover, we take computational issues into consideration. We especially theoretically investigate two types of optimization methods, robust gradient descent algorithm and empirical risk-based methods. With an extensive numerical study, we find that the optimizer based on empirical risks via Catoni-style estimation indeed shows better performance than other baselines. It indicates that estimation directly based on truncated data may lead to unsatisfactory results.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/01/2017

Efficient learning with robust gradient descent

Minimizing the empirical risk is a popular training strategy, but for le...
research
01/31/2022

Robust supervised learning with coordinate gradient descent

This paper considers the problem of supervised learning with linear meth...
research
10/15/2018

Robust descent using smoothed multiplicative noise

To improve the off-sample generalization of classical procedures minimiz...
research
01/27/2023

Robust variance-regularized risk minimization with concomitant scaling

Under losses which are potentially heavy-tailed, we consider the task of...
research
05/21/2018

Learning with Non-Convex Truncated Losses by SGD

Learning with a convex loss function has been a dominating paradigm for...
research
12/14/2020

Better scalability under potentially heavy-tailed feedback

We study scalable alternatives to robust gradient descent (RGD) techniqu...
research
01/12/2015

Scaling-up Empirical Risk Minimization: Optimization of Incomplete U-statistics

In a wide range of statistical learning problems such as ranking, cluste...

Please sign up or login with your details

Forgot password? Click here to reset