Publication:
Arabic word stemming algorithms and retrieval effectiveness

Date
2013
Authors
Sembok T.M.T.
Ata B.A.
Journal Title
Journal ISSN
Volume Title
Publisher
Research Projects
Organizational Units
Journal Issue
Abstract
Documents retrieval in Information Retrieval Systems (IRS) is generally about retrieving of relevant documents pertaining to information needs. The more the system able to understand the contents of documents the more effective will be the retrieval outcomes. But understanding of the contents is a very complex task. Conventional IRS applies algorithms that can only approximate the meaning of document contents through keywords approach using vector space model. Keywords may be unstemmed or stemmed. When keywords are stemmed and conflated in retrieval process, we are a step forwards in applying semantic technology in IRS. Word stemming is a process in morphological analysis under natural language processing, before syntactic and semantic analysis. We have developed algorithms for Arabic stemming and incorporated it in our experimental system in order to measure retrieval effectiveness. The results have shown that the retrieval effectiveness has increased when stemming is used.
Description
Keywords
Artificial intelligence , Information retrieval , Natural language processing , Algorithms , Artificial intelligence , Information retrieval systems , Natural language processing systems , Semantics , Vector spaces , Experimental system , Morphological analysis , NAtural language processing , Relevant documents , Retrieval effectiveness , Retrieval process , Semantic technologies , Vector space models , Information retrieval
Citation
Collections