Negative Confidence-Aware Weakly Supervised Binary Classification for Effective Review Helpfulness Classification

08/14/2020
by   Xi Wang, et al.
0

The incompleteness of positive labels and the presence of many unlabelled instances are common problems in binary classification applications such as in review helpfulness classification. Various studies from the classification literature consider all unlabelled instances as negative examples. However, a classification model that learns to classify binary instances with incomplete positive labels while assuming all unlabelled data to be negative examples will often generate a biased classifier. In this work, we propose a novel Negative Confidence-aware Weakly Supervised approach (NCWS), which customises a binary classification loss function by discriminating the unlabelled examples with different negative confidences during the classifier's training. We use the review helpfulness classification as a test case for examining the effectiveness of our NCWS approach. We thoroughly evaluate NCWS by using three different datasets, namely one from Yelp (venue reviews), and two from Amazon (Kindle and Electronics reviews). Our results show that NCWS outperforms strong baselines from the literature including an existing SVM-based approach (i.e. SVM-P), the positive and unlabelled learning-based approach (i.e. C-PU) and the positive confidence-based approach (i.e. P-conf) in addressing the classifier's bias problem. Moreover, we further examine the effectiveness of NCWS by using its classified helpful reviews in a state-of-the-art review-based venue recommendation model (i.e. DeepCoNN) and demonstrate the benefits of using NCWS in enhancing venue recommendation effectiveness in comparison to the baselines.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/29/2020

Binary Classification from Positive Data with Skewed Confidence

Positive-confidence (Pconf) classification [Ishida et al., 2018] is a pr...
research
08/02/2022

Binary Classification with Positive Labeling Sources

To create a large amount of training labels for machine learning models ...
research
03/11/2022

Classification from Positive and Biased Negative Data with Skewed Labeled Posterior Probability

The binary classification problem has a situation where only biased data...
research
08/13/2021

Adaptive Positive-Unlabelled Learning via Markov Diffusion

Positive-Unlabelled (PU) learning is the machine learning setting in whi...
research
12/15/2018

Weakly supervised segment annotation via expectation kernel density estimation

Since the labelling for the positive images/videos is ambiguous in weakl...
research
11/09/2022

Optimized Global Perturbation Attacks For Brain Tumour ROI Extraction From Binary Classification Models

Deep learning techniques have greatly benefited computer-aided diagnosti...
research
05/04/2017

Learning with Confident Examples: Rank Pruning for Robust Classification with Noisy Labels

Noisy PN learning is the problem of binary classification when training ...

Please sign up or login with your details

Forgot password? Click here to reset