Telugu Text Categorization using Language Models
Telugu Text Categorization using Language Models
Article PDF

Keywords

text categorization
language dependent and independent models
k-nearest neighbors

How to Cite

Swapna Narala, B. Padmaja Rani, & K. Ramakrishna. (2017). Telugu Text Categorization using Language Models. Global Journal of Computer Science and Technology, 16(H4), 9–14. Retrieved from https://gjcst.com/index.php/gjcst/article/view/733

Abstract

Document categorization has become an emerging technique in the field of research due to the abundance of documents available in digital form In this paper we propose language dependent and independent models applicable to categorization of Telugu documents India is a multilingual country a provision is made for each of the Indian states to choose their own authorized language for communicating at the state level for legitimate purpose The availability of constantly increasing amount of textual data of various Indian regional languages in electronic form has accelerated Hence the Classification of text documents based on languages is crucial Telugu is the third most spoken language in India and one of the fifteen most spoken language n the world It is the official language of the states of Telangana and Andhra Pradesh A variant of k-nearest neighbors algorithm used for categorization process The results obtained by the Comparisons of language dependent and independent models
Article PDF
Creative Commons License

This work is licensed under a Creative Commons Attribution 4.0 International License.

Copyright (c) 2016 Authors and Global Journals Private Limited