Qari OCR Arabic 0.2.2.1 VL 2B Instruct Model This is the model described in the paper QARI OCR: High Fidelity Arabic Text Recognition through Multimodal Large Language Model Adaptation. Model Overview This model is a fine tuned version of unsloth/Qwen2 VL 2B Instruct on an Arabic OCR dataset. It is optimized to perform Arabic Optical Character Recognition (OCR) for full page text. Code can be found at this url . Key Features Superior Accuracy : Achieves state of the art performance metrics for Arabic OCR Diacritics Support : Full recognition of Arabic diacritical marks (tashkeel) including fatḥah, kasrah, ḍammah, sukūn, shadda, and tanwin forms a strength confirmed by evaluation on a primarily diacritical text dataset Multiple Font Support : Works across a variety of Arabic font styles Layout Flexibility : Handles different document layouts and formats Model Details Base Model : Qwen2 VL Fine tuning Dataset : Arabic OCR dataset Objective : Extract full page Arabic text with high accuracy Languages : Arabic Tasks : OCR (Optical Character Recognition) Dataset size : 50,000 records Epochs : 1 Evaluation Metrics Performance is evaluated using three standard metrics: Word Error Rate (WE…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy