SanskritOCR

SanskritOCR is an OCR in Indian Language for Sanskrit, Hindi and other Indian languages based on Devanagari script.

Sanskrit OCR is developed by a Sanskrit scholar from Germany - Dr. Oliver Hellwig of Department for Languages and Cultures of Southern Asia, Freie Universität Berlin. The official website is in German. The interface of earlier versions of the software was also in German, but later versions have an English interface too.[1][2][3]

The software scans the printed Devanagari text and converts it into Romananized text. The romanized text than can be converted to Devanagari using this online converter.

Dr. Hellwig is currently developing a specialized Hindi OCR, which has recognition rates of about 99.5% on good documents and hope to make this program available by the end of the year.[4]

References

  1. Prabhu, S. (2020-06-04). "Pazhur Patasala — a revival story". The Hindu. ISSN 0971-751X. Retrieved 2021-09-01. An OCR (Optical Character Recognition) for Sanskrit has created an offline corpus that includes over 3,000 books.
  2. "Digitisation going on at brisk pace: Vice-Chancellor Prof V Muralidhara Sharma". www.thehansindia.com. Hans News Service. 2019-03-20. Retrieved 2021-09-01.
  3. Dikshit, Ashish (2016-10-27). "Who Says Sanskrit Is Dead? It's Rocking the Wiki World". TheQuint. Retrieved 2021-09-01.
  4. "image processing - OCR for Devanagari (Hindi / Marathi / Sanskrit)". Stack Overflow. Retrieved 2021-09-01.


This article is issued from Wikipedia. The text is licensed under Creative Commons - Attribution - Sharealike. Additional terms may apply for the media files.