Active Learning for Regression with Aggregated Outputs

10/04/2022
by   Tomoharu Iwata, et al.
0

Due to the privacy protection or the difficulty of data collection, we cannot observe individual outputs for each instance, but we can observe aggregated outputs that are summed over multiple instances in a set in some real-world applications. To reduce the labeling cost for training regression models for such aggregated data, we propose an active learning method that sequentially selects sets to be labeled to improve the predictive performance with fewer labeled sets. For the selection measurement, the proposed method uses the mutual information, which quantifies the reduction of the uncertainty of the model parameters by observing the aggregated output. With Bayesian linear basis functions for modeling outputs given an input, which include approximated Gaussian processes and neural networks, we can efficiently calculate the mutual information in a closed form. With the experiments using various datasets, we demonstrate that the proposed method achieves better predictive performance with fewer labeled sets than existing methods.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/30/2021

BABA: Beta Approximation for Bayesian Active Learning

This paper introduces a new acquisition function under the Bayesian acti...
research
07/19/2017

Improving Output Uncertainty Estimation and Generalization in Deep Learning via Neural Network Gaussian Processes

We propose a simple method that combines neural networks and Gaussian pr...
research
01/01/2021

Active Learning Under Malicious Mislabeling and Poisoning Attacks

Deep neural networks usually require large labeled datasets for training...
research
06/20/2012

On Discarding, Caching, and Recalling Samples in Active Learning

We address challenges of active learning under scarce informational reso...
research
04/17/2023

Prediction-Oriented Bayesian Active Learning

Information-theoretic approaches to active learning have traditionally f...
research
10/06/2017

Bag-Level Aggregation for Multiple Instance Active Learning in Instance Classification Problems

A growing number of applications, e.g. video surveillance and medical im...
research
05/04/2021

Multi-resolution Spatial Regression for Aggregated Data with an Application to Crop Yield Prediction

We develop a new methodology for spatial regression of aggregated output...

Please sign up or login with your details

Forgot password? Click here to reset