Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Effectiveness of Human-in-the-Loop Sentiment Polarization with Few Corrected Labels

Author 1: Ruhaila Maskat Author 2: Nurzety Aqtar Ahmad Azuan Author 3: Siti Auni Amaram Author 4: Nur Hayatin
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 13, No. 7 · Published 2022

DOI: https://doi.org/10.14569/IJACSA.2022.0130726

Abstract

In this work, we investigated the effectiveness of adopting Human-in-the-Loop (HITL) aimed to correct automatically generated labels from existing scoring models, e.g. SentiWordNet and Vader to enhance prediction accuracy. Recently, many proposals showed a trend in utilizing these models to label data by assuming that the labels produced are near to ground truth. However, none investigated the correctness of this notion. Therefore, this paper fills this gap. Bad labels result in bad predictions, hence hypothetically, by positioning a human in the computing loop to correct inaccurate labels accuracy performance can be improved. As it is infeasible to expect a human to correct a multitude of labels, we set out to answer the questions of “What is the smallest percentage of corrected labels needed to improve prediction quality against a baseline?” and “Would randomly selecting automatic labels for correction produce better prediction than specifically choosing labels with distinct data points?”. Naïve Bayes (NB) and Decision Tree (DT) were employed on AirBnB and Vaccines public datasets. We could conclude from our results that not all ML algorithms are suited to be used in a HITL environment. NB fared better than DT at producing improved accuracy with small percentages of corrected labels, as low as 1%, exceeding the baseline. When selected for human correction, labels with distinct data points assisted in enhancing the accuracy better than random selection for NB across both datasets, yet partially for DT.

Keywords

How to Cite this Article

Maskat, R., Azuan, N. A. A., Amaram, S. A., & Hayatin, N. (2022). Effectiveness of Human-in-the-Loop Sentiment Polarization with Few Corrected Labels. International Journal of Advanced Computer Science and Applications, 13(7). https://doi.org/10.14569/IJACSA.2022.0130726

Maskat, Ruhaila, et al.. "Effectiveness of Human-in-the-Loop Sentiment Polarization with Few Corrected Labels." International Journal of Advanced Computer Science and Applications, vol. 13, no. 7, 2022, https://doi.org/10.14569/IJACSA.2022.0130726.

@article{Maskat2022,
  title     = {Effectiveness of Human-in-the-Loop Sentiment Polarization with Few Corrected Labels},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {13},
  number    = {7},
  year      = {2022},
  publisher = {The Science and Information Organization},
  author    = {Ruhaila Maskat and Nurzety Aqtar Ahmad Azuan and Siti Auni Amaram and Nur Hayatin},
  doi       = {10.14569/IJACSA.2022.0130726},
  url       = {https://doi.org/10.14569/IJACSA.2022.0130726}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.