Arabic-first Document Intelligence API platform for OCR, Arabic handwriting recognition, document extraction, validation, search, chat with documents, and structured JSON output.
-
Updated
Oct 5, 2026 - Python
Arabic-first Document Intelligence API platform for OCR, Arabic handwriting recognition, document extraction, validation, search, chat with documents, and structured JSON output.
This research aims to fine-tune an Arabic OCR model using Tesseract 5.0, enhancing text recognition accuracy through extensive data collection, preprocessing, and image generation. By leveraging advanced training techniques and data augmentation, we achieve significant improvements in word error rates (WER).
In this repository, OCR-related datasets are available.
Arabic Chat with PDF is a user-friendly application that lets you interact with Arabic PDF documents. Powered by advanced language models, OCR, and vector search, it allows you to upload PDFs, ask questions, and receive accurate Arabic responses 🚀
الدليل الحي والمفتوح للبرمجيات والنماذج مفتوحة المصدر للغة العربية | The centralized living directory & showcase for Arabic open-source software, NLP, models, and tools.
Alef-OCR-Image2Html, an OCR model designed to transform Arabic documents including historical texts, scanned pages, and handwritten materials into structured and semantic HTML.
Automatic license-plate recognition for Sudanese plates — YOLO detector + fine-tuned OCR, with a reproducible benchmark.
The largest publicly released line-level dataset of historical Arabic manuscripts — 14 books, 3,043 pages, 28,600 lines, with margin/insertion-anchor annotations for non-linear reading order.
Official code for "Ketaba-OCR at AR-MS NakbaNLP 2026" — QLoRA fine-tuning of a specialized HTR model with Linear+Boost ensemble for Arabic manuscript recognition. 1st place per-line (CER 0.082) and 3rd place official leaderboard at NakbaNLP 2026 (LREC 2026).
Optical Character Recognition, OCR pipeline, Arabic OCR, Deep Learning OCR, Computer Vision text extraction, Text recognition system, AI document processing, Multilingual OCR, Transformer OCR, OCR benchmarking, Bounding box detection, Ground truth evaluation.
Fast, zero-dependency screen OCR to clipboard for Linux (Wayland & X11). Auto multi-language (Arabic + English).
Arabic handwritten text recognition using a CRNN (CNN + BiLSTM) with CTC loss, trained on the KHATT dataset — includes a Gradio web demo for OCR on your own images.
An AI-powered OCR and document processing system designed to convert Arabic PDF books and images into high-quality, editable scientific text layouts
Additional experimental model for NakbaNLP 2026 Shared Task (AR-MS) — LoRA/DoRA fine-tuning of Qari-OCR (Qwen2-VL-2B) for Arabic handwritten manuscript recognition on the Omar Al-Saleh Memoir Collection (1951-1965).
DocuVision Electron — Arabic/English OCR desktop app for local document scanning, review, search, and export on Windows and Linux.
DocuVision Showcase — commercial Arabic/English OCR desktop application for document review, search, and export; source code is not included.
Nassij V3: High-accuracy Arabic PDF-to-DOCX converter with direct digital extraction (NassijScanner) and cryptographic linguistic integrity verification (Merkle proofs).
Reproducible Arabic OCR benchmark by Clouda OCR — distorted documents, normalized Arabic CER, model comparisons, methodology, and public audit artifacts.
Arabic OCR and diacritization pipeline with FastAPI backend, Android client, and reproducible notebook workflow.
Image-to-text system for recognizing Arabic handwritten manuscripts using image feature extraction and a Transformer decoder.
To associate your repository with the arabic-ocr topic, visit your repo's landing page and select "manage topics."