Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Comparative Analysis of Lexicon and Machine Learning Approach for Sentiment Analysis

Author 1: Roopam Srivastava Author 2: P. K. Bharti Author 3: Parul Verma
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 13, No. 3 · Published 2022 · Cited by 28

DOI: https://doi.org/10.14569/IJACSA.2022.0130312

Abstract

Opinion mining or analysis of text are other terms for sentiment analysis. The fundamental objective is to extract meaningful information and data from unstructured text using natural language processing, statistical, and linguistics methodologies. This further is used for deriving qualitative and quantitative results on the scale of ‘positive’, ‘neutral’, or ‘negative to get the overall sentiment analysis. In this research, we worked with both approaches, machine learning, and an unsupervised lexicon-based algorithm for sentiment calculation and model performance. Stochastic gradient descent (SGD) is utilized in this work for optimization for support vector machine (SVM) and logistic regression. AFINN and Vader lexicon are used for the lexicon model. Both the feature TF-IDF and bag of a word are used for classification. This dataset includes "Trip advisor hotel reviews". There are around 20k reviews in the dataset. Cleaned and preprocessed data were used in our work. We conducted some training and assessment. A classifier's accuracy is measured using evaluation metrics. In TF-IDF, the Support Vector Machine is the more accurate of the two classifiers used to assess machine learning accuracy. The classification rate in Bag of Words was 95.2 percent and the accuracy in TF-IDF was 96.3 percent on the support vector machine algorithm. VADER outperforms the Lexicon model with an accuracy of 88.7%, whereas AFINN Lexicon has an accuracy of 86.0%. When comparing the Supervised and unsupervised lexicon approaches, support vector machine model outperforms with a TFIDF accuracy of 96.3 percent and a VADER lexicon accuracy of 88.7%.

Keywords

How to Cite this Article

Srivastava, R., Bharti, P. K., & Verma, P. (2022). Comparative Analysis of Lexicon and Machine Learning Approach for Sentiment Analysis. International Journal of Advanced Computer Science and Applications, 13(3). https://doi.org/10.14569/IJACSA.2022.0130312

Srivastava, Roopam, et al.. "Comparative Analysis of Lexicon and Machine Learning Approach for Sentiment Analysis." International Journal of Advanced Computer Science and Applications, vol. 13, no. 3, 2022, https://doi.org/10.14569/IJACSA.2022.0130312.

@article{Srivastava2022,
  title     = {Comparative Analysis of Lexicon and Machine Learning Approach for Sentiment Analysis},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {13},
  number    = {3},
  year      = {2022},
  publisher = {The Science and Information Organization},
  author    = {Roopam Srivastava and P. K. Bharti and Parul Verma},
  doi       = {10.14569/IJACSA.2022.0130312},
  url       = {https://doi.org/10.14569/IJACSA.2022.0130312}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.