Computer Science – Computation and Language
Scientific paper
2011-11-10
International Journal of Computer Science & Information Technology (IJCSIT) Vol 3, No 5, Oct 2011, pp 53-66
Computer Science
Computation and Language
14 pages, 6 figures, see http://airccse.org/journal/jcsit/1011csit05.pdf
Scientific paper
10.5121/ijcsit.2011.3505
This paper deals with the identification of Multiword Expressions (MWEs) in Manipuri, a highly agglutinative Indian Language. Manipuri is listed in the Eight Schedule of Indian Constitution. MWE plays an important role in the applications of Natural Language Processing(NLP) like Machine Translation, Part of Speech tagging, Information Retrieval, Question Answering etc. Feature selection is an important factor in the recognition of Manipuri MWEs using Conditional Random Field (CRF). The disadvantage of manual selection and choosing of the appropriate features for running CRF motivates us to think of Genetic Algorithm (GA). Using GA we are able to find the optimal features to run the CRF. We have tried with fifty generations in feature selection along with three fold cross validation as fitness function. This model demonstrated the Recall (R) of 64.08%, Precision (P) of 86.84% and F-measure (F) of 73.74%, showing an improvement over the CRF based Manipuri MWE identification without GA application.
Bandyopadhyay Sivaji
Nongmeikapam Kishorjit
No associations
LandOfFree
Genetic Algorithm (GA) in Feature Selection for CRF Based Manipuri Multiword Expression (MWE) Identification does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.
If you have personal experience with Genetic Algorithm (GA) in Feature Selection for CRF Based Manipuri Multiword Expression (MWE) Identification, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Genetic Algorithm (GA) in Feature Selection for CRF Based Manipuri Multiword Expression (MWE) Identification will most certainly appreciate the feedback.
Profile ID: LFWR-SCP-O-727689