A DEEP LEARNING-BASED APPROACH FOR URDU AND ENGLISH HANDWRITTEN TEXT RECOGNITION
Keywords:
Transformer-based models, multilingual, Optical character recognition (OCR), Urdu OCR, document analysis, and performance evaluationAbstract
The difficulties presented by low-resource languages are significant when it comes to digitizing written material. Due to the simple linguistic resources that mostly abound within such languages, establishing efficient systems for accurate optical character recognition is an issue that calls for special consideration. Asides highlighting the importance and need to specialize on such languages, this paper introduces ViLanOCR, which is an innovative multilingual OCR system tailored specifically towards Urdu and English languages. The system takes advantage of sophisticated multilingual transformers within its language models that result in greater performances than other approaches, which pose considerable difficulties due to the nature of low resource languages. On the Urdu UHWR testing set, the proposed system achieves state-of-the-art performances with a CER measure of 1.1%. The testing findings show how effective the suggested method is, outperforming Cutting-edge beginnings in the digitization of Urdu handwriting












