Adapting Multilingual Vision Language Transformers for Low-Resource Urdu Optical Character Recognition (OCR).
| aut.relation.articlenumber | e1964 | |
| aut.relation.journal | PeerJ Comput Sci | |
| aut.relation.startpage | e1964 | |
| aut.relation.volume | 10 | |
| dc.contributor.author | Cheema, Musa Dildar Ahmed | |
| dc.contributor.author | Shaiq, Mohammad Daniyal | |
| dc.contributor.author | Mirza, Farhaan | |
| dc.contributor.author | Kamal, Ali | |
| dc.contributor.author | Naeem, M Asif | |
| dc.date.accessioned | 2024-05-16T23:24:10Z | |
| dc.date.available | 2024-05-16T23:24:10Z | |
| dc.date.issued | 2024-04-29 | |
| dc.description.abstract | In the realm of digitizing written content, the challenges posed by low-resource languages are noteworthy. These languages, often lacking in comprehensive linguistic resources, require specialized attention to develop robust systems for accurate optical character recognition (OCR). This article addresses the significance of focusing on such languages and introduces ViLanOCR, an innovative bilingual OCR system tailored for Urdu and English. Unlike existing systems, which struggle with the intricacies of low-resource languages, ViLanOCR leverages advanced multilingual transformer-based language models to achieve superior performances. The proposed approach is evaluated using the character error rate (CER) metric and achieves state-of-the-art results on the Urdu UHWR dataset, with a CER of 1.1%. The experimental results demonstrate the effectiveness of the proposed approach, surpassing state of the-art baselines in Urdu handwriting digitization. | |
| dc.identifier.citation | PeerJ Comput Sci, ISSN: 2167-9843 (Print); 2376-5992 (Online), PeerJ, 10, e1964-. doi: 10.7717/peerj-cs.1964 | |
| dc.identifier.doi | 10.7717/peerj-cs.1964 | |
| dc.identifier.issn | 2167-9843 | |
| dc.identifier.issn | 2376-5992 | |
| dc.identifier.uri | http://hdl.handle.net/10292/17559 | |
| dc.language | eng | |
| dc.publisher | PeerJ | |
| dc.relation.uri | https://peerj.com/articles/cs-1964/ | |
| dc.rights.accessrights | OpenAccess | |
| dc.rights.uri | http://www.creativecommons.org/licenses/by/4.0/ | |
| dc.subject | Document analysis | |
| dc.subject | Multilingual | |
| dc.subject | OCR | |
| dc.subject | Performance evaluation | |
| dc.subject | Transformer based models | |
| dc.subject | Urdu OCR | |
| dc.subject | 46 Information and Computing Sciences | |
| dc.subject | 0806 Information Systems | |
| dc.subject | 46 Information and computing sciences | |
| dc.title | Adapting Multilingual Vision Language Transformers for Low-Resource Urdu Optical Character Recognition (OCR). | |
| dc.type | Journal Article | |
| pubs.elements-id | 547979 |
Files
Original bundle
1 - 1 of 1
Loading...
- Name:
- Adapting multilingual vision language transformers for low-resource Urdu optical character recognition (OCR).pdf
- Size:
- 7.96 MB
- Format:
- Adobe Portable Document Format
- Description:
- Journal article
