• DocumentCode
    2657333
  • Title

    Softconverter: A novel approach to construct OCR for printed Urdu isolated characters

  • Author

    Tariq, Junaid ; Nauman, Umar ; Naru, Muhammad Umair

  • Author_Institution
    Dept. of Comput. Sci., COMSATS Inst. of Inf. Technol., Islamabad, Pakistan
  • Volume
    3
  • fYear
    2010
  • fDate
    16-18 April 2010
  • Abstract
    Urdu covers the large part of sub-continent´s (Pakistan, India, & Bangladesh) literature, which is present in hard form. It is difficult to change, distribute information present in hard form than soft form. Soft form can be uploaded on internet, and can be edited and reprinted. The problem of converting books containing Urdu characters to soft form can be done with Urdu OCR (Optical Character Recognition). NN (Neural network) is used to constructs OCR, but it makes the development of OCR very difficult and complex, even if the font size and font style is fixed. In this paper we present a simple and easy way to construct OCR for isolated characters of Urdu language (or right to left writing) called “softconverter” with the help of database, without using neural network. This paper proves that OCR can be implemented without using NN. Our prototype of Softconverter has accuracy rate of 97.43%.
  • Keywords
    natural languages; neural nets; optical character recognition; OCR; Urdu language; neural network; optical character recognition; printed Urdu isolated characters; softconverter; Character recognition; Computer science; Image resolution; Information technology; Natural languages; Neural networks; Optical character recognition software; Optical computing; Pixel; Prototypes;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Computer Engineering and Technology (ICCET), 2010 2nd International Conference on
  • Conference_Location
    Chengdu
  • Print_ISBN
    978-1-4244-6347-3
  • Type

    conf

  • DOI
    10.1109/ICCET.2010.5485836
  • Filename
    5485836