Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Automatic Keyphrase Extractor from Arabic Documents

Author 1: Hassan M. Najadat Author 2: Ismail I. Hmeidi Author 3: Mohammed N. Al-Kabi Author 4: Maysa Mahmoud Bany Issa
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 7, No. 2 · Published 2016 · Cited by 12

DOI: https://doi.org/10.14569/IJACSA.2016.070226

Abstract

The keyphrase is a sentence or a part of a sentence that contains a sequence of words that expresses the meaning and the purpose of any given paragraph. Keyphrase extraction is the task of identifying the possible keyphrases from a given document. Many applications including text summarization, indexing, and characterization use keyphrase extraction. Also, it is an essential task to improve the performance of any information retrieval system. The internet contains a massive amount of documents that may have been manually assigned keyphrases or not. The Arabic language is an important language in the world. Nowadays the number of online Arabic documents is growing rapidly; and most of them have no manually assigned keyphrases, so the user will scan the whole retrieved web documents. To avoid scanning the entire retrieved document, we need keyphrases assigned to each web document manually or automatically. This paper addresses the problem of identifying keyphrases in Arabic documents automatically. In this work, we provide a novel algorithm that identified keyphrases from Arabic text. The new algorithm, Automatic Keyphrases Extraction from Arabic (AKEA), extracts keyphrases from Arabic documents automatically. In order to test the algorithm, we collected a dataset containing 100 documents from Arabic wiki; also, we downloaded another 56 agricultural documents from Food and Agricultural Organization of the United Nations (F.A.O.). The evaluation results show that the system achieves 83% precision value in identifying 2-word and 3-word keyphrases from agricultural domains.

Keywords

How to Cite this Article

Najadat, H. M., Hmeidi, I. I., Al-Kabi, M. N., & Issa, M. M. B. (2016). Automatic Keyphrase Extractor from Arabic Documents. International Journal of Advanced Computer Science and Applications, 7(2). https://doi.org/10.14569/IJACSA.2016.070226

Najadat, Hassan M., et al.. "Automatic Keyphrase Extractor from Arabic Documents." International Journal of Advanced Computer Science and Applications, vol. 7, no. 2, 2016, https://doi.org/10.14569/IJACSA.2016.070226.

@article{Najadat2016,
  title     = {Automatic Keyphrase Extractor from Arabic Documents},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {7},
  number    = {2},
  year      = {2016},
  publisher = {The Science and Information Organization},
  author    = {Hassan M. Najadat and Ismail I. Hmeidi and Mohammed N. Al-Kabi and Maysa Mahmoud Bany Issa},
  doi       = {10.14569/IJACSA.2016.070226},
  url       = {https://doi.org/10.14569/IJACSA.2016.070226}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.