An Implicit Segmentation Approach for Telugu Text Recognition Based on Hidden Markov Models

Autor:	Atul Negi, D. Koteswara Rao
Rok vydání:	2015
Předmět:	Computer science business.industry Bigram Feature vector Speech recognition computer.software_genre Telugu language.human_language ComputingMethodologies_PATTERNRECOGNITION Word recognition language Segmentation Language model Artificial intelligence business Hidden Markov model computer Natural language processing Word (computer architecture)
Zdroj:	Advances in Intelligent Systems and Computing ISBN: 9783319286563 SIRS
DOI:	10.1007/978-3-319-28658-7_54
Popis:	Telugu text is composed of aksharas (characters). The presence of split and connected aksharas in Telugu document images causes segmentation difficulties and the performance of the Telugu OCR systems is affected. Our novel approach to solve this problem is using an implicit segmentation for recognizing words. The implicit segmentation approach does not need prior segmentation of the words into aksharas before they are recognized. Since the Hidden Markov models (HMM) are successfully applied for phoneme recognition with no prior segmentation of the speech into phonemes in the automatic speech recognition applications. In this paper, we report on the use of continuous density Hidden Markov Models for representing the shape of aksharas to build Telugu text recognition system. The sliding window method is used for computing simple statistical features and 450 akshara HMMs are trained. We use word bigram language model as contextual information. The word recognition relies on akshara models and contextual information of words. The word recognition involves finding the maximum likelihood sequence of akshara models that matches against the feature vector sequence. Our system recognizes words with split and connected aksharas. The performance of the system is encouraging.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_________::010daff9da3bd296d1b3d16383fd2595 https://doi.org/10.1007/978-3-319-28658-7_54 Zobrazit plný text záznamu