Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Improving Arabic Cognitive Distortion Classification in Twitter using BERTopic

Author 1: Fatima Alhaj Author 2: Ali Al-Haj Author 3: Ahmad Sharieh Author 4: Riad Jabri
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 13, No. 1 · Published 2022 · Cited by 56

DOI: https://doi.org/10.14569/IJACSA.2022.0130199

Abstract

Social media platforms allow users to share thoughts, experiences, and beliefs. These platforms represent a rich resource for natural language processing techniques to make inferences in the context of cognitive psychology. Some inaccurate and biased thinking patterns are defined as cognitive distortions. Detecting these distortions helps users restructure how to perceive thoughts in a healthier way. This paper proposed a machine learning-based approach to improve cognitive distortions’ classi-fication of the Arabic content over Twitter. One of the challenges that face this task is the text shortness, which results in a sparsity of co-occurrence patterns and a lack of context information (semantic features). The proposed approach enriches text rep-resentation by defining the latent topics within tweets. Although classification is a supervised learning concept, the enrichment step uses unsupervised learning. The proposed algorithm utilizes a transformer-based topic modeling (BERTopic). It employs two types of document representations and performs averaging and concatenation to produce contextual topic embeddings. A comparative analysis of F1-score, precision, recall, and accuracy is presented. The experimental results demonstrate that our enriched representation outperformed the baseline models by different rates. These encouraging results suggest that using latent topic distribution, obtained from the BERTopic technique, can improve the classifier’s ability to distinguish between different CD categories.

Keywords

How to Cite this Article

Alhaj, F., Al-Haj, A., Sharieh, A., & Jabri, R. (2022). Improving Arabic Cognitive Distortion Classification in Twitter using BERTopic. International Journal of Advanced Computer Science and Applications, 13(1). https://doi.org/10.14569/IJACSA.2022.0130199

Alhaj, Fatima, et al.. "Improving Arabic Cognitive Distortion Classification in Twitter using BERTopic." International Journal of Advanced Computer Science and Applications, vol. 13, no. 1, 2022, https://doi.org/10.14569/IJACSA.2022.0130199.

@article{Alhaj2022,
  title     = {Improving Arabic Cognitive Distortion Classification in Twitter using BERTopic},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {13},
  number    = {1},
  year      = {2022},
  publisher = {The Science and Information Organization},
  author    = {Fatima Alhaj and Ali Al-Haj and Ahmad Sharieh and Riad Jabri},
  doi       = {10.14569/IJACSA.2022.0130199},
  url       = {https://doi.org/10.14569/IJACSA.2022.0130199}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.