Morphological Reconstruction for Word Level Script Identification

Computer Science – Computer Vision and Pattern Recognition

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

11 Pages, 8 Figures,5 Tables; Revised: 15-06-2007,Published: 30-06-2007

Scientific paper

A line of a bilingual document page may contain text words in regional language and numerals in English. For Optical Character Recognition (OCR) of such a document page, it is necessary to identify different script forms before running an individual OCR system. In this paper, we have identified a tool of morphological opening by reconstruction of an image in different directions and regional descriptors for script identification at word level, based on the observation that every text has a distinct visual appearance. The proposed system is developed for three Indian major bilingual documents, Kannada, Telugu and Devnagari containing English numerals. The nearest neighbour and k-nearest neighbour algorithms are applied to classify new word images. The proposed algorithm is tested on 2625 words with various font styles and sizes. The results obtained are quite encouraging

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

Morphological Reconstruction for Word Level Script Identification does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with Morphological Reconstruction for Word Level Script Identification, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Morphological Reconstruction for Word Level Script Identification will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-145010

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.