首页    期刊浏览 2025年07月16日 星期三
登录注册

文章基本信息

  • 标题:Automatic Keyphrase Extractor from Arabic Documents
  • 本地全文:下载
  • 作者:Hassan M. Najadat ; Ismail I. Hmeidi ; Mohammed N. Al-Kabi
  • 期刊名称:International Journal of Advanced Computer Science and Applications(IJACSA)
  • 印刷版ISSN:2158-107X
  • 电子版ISSN:2156-5570
  • 出版年度:2016
  • 卷号:7
  • 期号:2
  • DOI:10.14569/IJACSA.2016.070226
  • 出版社:Science and Information Society (SAI)
  • 摘要:The keyphrase is a sentence or a part of a sentence that contains a sequence of words that expresses the meaning and the purpose of any given paragraph. Keyphrase extraction is the task of identifying the possible keyphrases from a given document. Many applications including text summarization, indexing, and characterization use keyphrase extraction. Also, it is an essential task to improve the performance of any information retrieval system. The internet contains a massive amount of documents that may have been manually assigned keyphrases or not. The Arabic language is an important language in the world. Nowadays the number of online Arabic documents is growing rapidly; and most of them have no manually assigned keyphrases, so the user will scan the whole retrieved web documents. To avoid scanning the entire retrieved document, we need keyphrases assigned to each web document manually or automatically. This paper addresses the problem of identifying keyphrases in Arabic documents automatically. In this work, we provide a novel algorithm that identified keyphrases from Arabic text. The new algorithm, Automatic Keyphrases Extraction from Arabic (AKEA), extracts keyphrases from Arabic documents automatically. In order to test the algorithm, we collected a dataset containing 100 documents from Arabic wiki; also, we downloaded another 56 agricultural documents from Food and Agricultural Organization of the United Nations (F.A.O.). The evaluation results show that the system achieves 83% precision value in identifying 2-word and 3-word keyphrases from agricultural domains.
  • 关键词:thesai; IJACSA; thesai.org; journal; IJACSA papers; Arabic Keyphrase Extraction; Unsupervised Arabic Keyphrase Extraction; Information Retrieval
国家哲学社会科学文献中心版权所有