International Journal of Computer Applications |
Foundation of Computer Science (FCS), NY, USA |
Volume 51 - Number 14 |
Year of Publication: 2012 |
Authors: Kh Raju Singha, Bipul Syam Purkayastha, Kh Dhiren Singha |
10.5120/8111-1727 |
Kh Raju Singha, Bipul Syam Purkayastha, Kh Dhiren Singha . Part of Speech Tagging in Manipuri: A Rule based Approach. International Journal of Computer Applications. 51, 14 ( August 2012), 31-36. DOI=10.5120/8111-1727
The process of assigning morpho-syntactic categories of each morpheme including punctuation marks in a given text document according to the context is called Part of Speech (POS) tagging. In this paper we represent the rule-based Part of Speech Tagger of Manipuri by applying a set of hand written linguistic rules of Manipuri language. Nevertheless, it is very difficult to classify the lexical categories of Manipuri, an agglutinating Tibeto-Burman language of Northeast India. So, in this tagger we are using the affix stripping technique to segment the affixes from the root. As Manipuri has limited POS tagged corpus, the tagged output of this tagger will be very helpful to analyze Manipuri Part of speech by using many statistical models.