Efficient Feature Selection techniques for Sentiment Analysis

11/01/2019
by   Avinash Madasu, et al.
0

Sentiment analysis is a domain of study that focuses on identifying and classifying the ideas expressed in the form of text into positive, negative and neutral polarities. Feature selection is a crucial process in machine learning. In this paper, we aim to study the performance of different feature selection techniques for sentiment analysis. Term Frequency Inverse Document Frequency (TF-IDF) is used as the feature extraction technique for creating feature vocabulary. Various Feature Selection (FS) techniques are experimented to select the best set of features from feature vocabulary. The selected features are trained using different machine learning classifiers Logistic Regression (LR), Support Vector Machines (SVM), Decision Tree (DT) and Naive Bayes (NB). Ensemble techniques Bagging and Random Subspace are applied on classifiers to enhance the performance on sentiment analysis. We show that, when the best FS techniques are trained using ensemble methods achieve remarkable results on sentiment analysis. We also compare the performance of FS methods trained using Bagging, Random Subspace with varied neural network architectures. We show that FS techniques trained using ensemble classifiers outperform neural networks requiring significantly less training time and parameters thereby eliminating the need for extensive hyper-parameter tuning.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/04/2019

A Study of Feature Extraction techniques for Sentiment Analysis

Sentiment Analysis refers to the study of systematically extracting the ...
research
05/03/2021

A Machine Learning Based Ensemble Method for Automatic Multiclass Classification of Decisions

Stakeholders make various types of decisions with respect to requirement...
research
05/02/2018

Automatic Coding for Neonatal Jaundice From Free Text Data Using Ensemble Methods

This study explores the creation of a machine learning model to automati...
research
01/27/2017

A Comparative Study on Different Types of Approaches to Bengali document Categorization

Document categorization is a technique where the category of a document ...
research
02/01/2021

Student sentiment Analysis Using Classification With Feature Extraction Techniques

Technical growths have empowered, numerous revolutions in the educationa...
research
12/27/2014

Persian Sentiment Analyzer: A Framework based on a Novel Feature Selection Method

In the recent decade, with the enormous growth of digital content in int...
research
04/01/2017

Sentiment Analysis of Citations Using Word2vec

Citation sentiment analysis is an important task in scientific paper ana...

Please sign up or login with your details

Forgot password? Click here to reset