{"title":"Named entity recognition in Assamese using CRFS and rules","authors":"Padmaja Sharma, U. Sharma, J. Kalita","doi":"10.1109/IALP.2014.6973498","DOIUrl":null,"url":null,"abstract":"Named Entity Recognition (NER) is an important task in all Natural Language Processing (NLP) applications. It is the process of identifying and classifying the proper noun into classes such as person, location, organization and miscellaneous. Substantial work has been done in English and other European languages, achieving greater accuracy compared to the Indian Languages. Although NER in Indian languages is a difficult and challenging task and suffers from scarcity of resources, such work has started to appear recently. This paper discusses work on NER in Assamese using both Conditional Random Fields and a Rule-Based approach which gives an F-measure of 90-95% accuracy.","PeriodicalId":117334,"journal":{"name":"2014 International Conference on Asian Language Processing (IALP)","volume":"10 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2014-10-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"9","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"2014 International Conference on Asian Language Processing (IALP)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/IALP.2014.6973498","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
引用次数: 9
Abstract
Named Entity Recognition (NER) is an important task in all Natural Language Processing (NLP) applications. It is the process of identifying and classifying the proper noun into classes such as person, location, organization and miscellaneous. Substantial work has been done in English and other European languages, achieving greater accuracy compared to the Indian Languages. Although NER in Indian languages is a difficult and challenging task and suffers from scarcity of resources, such work has started to appear recently. This paper discusses work on NER in Assamese using both Conditional Random Fields and a Rule-Based approach which gives an F-measure of 90-95% accuracy.